Team Ai
Datasetpublic

HumynLabs/Spanish_Documents_Dataset_PDF

Spanish Documents Dataset (PDF) This dataset contains a curated collection of Spanish-language documents in PDF format. It includes books, educational materials, research papers, news articles, and government publications written in Spanish. The dataset supports AI research in OCR, multilingual document understanding, and text extraction for Latin-script languages. Contact For queries or collaborations related to this dataset, contact: anoushka@kgen.io… See the full description on the dataset page: https://huggingface.co/datasets/HumynLabs/Spanish_Documents_Dataset_PDF.

sourceHugging Facecc-by-4.0updated 11mo agoView on Hugging Face
0likes721downloads
4 commits on main
0f33e6111mo ago

Update README.md

KAI
c0cc56d11mo ago

Upload 21 files

KAI
ad4924d11mo ago

Update README.md

KAI
3d9599a11mo ago

initial commit

KAI