Team Ai
Datasetpublic

SpringRollMonster/CIVQA-TesseractOCR-LayoutLM

CIVQA TesseractOCR LayoutLM Dataset The Czech Invoice Visual Question Answering dataset was created with Tesseract OCR and encoded for the LayoutLM. The pre-encoded dataset can be found on this link: https://huggingface.co/datasets/fimu-docproc-research/CIVQA-TesseractOCR All invoices used in this dataset were obtained from public sources. Over these invoices, we were focusing on 15 different entities, which are crucial for processing the invoices. Invoice number Variable… See the full description on the dataset page: https://huggingface.co/datasets/SpringRollMonster/CIVQA-TesseractOCR-LayoutLM.

sourceHugging Facemitupdated 2mo agoView on Hugging Face
0likes125downloads
1 commits on main
4f803dc2mo ago

Duplicate from fimu-docproc-research/CIVQA-TesseractOCR-LayoutLM

SpringRollMonster, Sharka