Dataset created with PDF2Dataset -- OCR + structure-aware chunking pipeline.
text
string
Raw markdown chunk with image refs… See the full description on the dataset page:
https://huggingface.co/datasets/Svngoku/Africans-THE-HISTORY-OF-A-CONTINENT-Second-Edition.