This dataset contains a curated collection of Spanish-language documents in PDF format. It includes books, educational materials, research papers, news articles, and government publications written in Spanish. The dataset supports AI research in OCR, multilingual document understanding, and text extraction for Latin-script languages.
Contact
For queries or collaborations related to this dataset, contact: