This dataset contains the full text content of Islamic Arabic books from the Shamela Library, organized by category, book, volume, and page, with footnotes stored separately. It is designed to support Arabic NLP, digital humanities, and bibliographic analysis.
🔗 This dataset is linked to the companion metadata dataset:
👉 Shamela_Books_info via the book_id field.
The dataset includes the original raw files as well as a single… See the full description on the dataset page:
https://huggingface.co/datasets/MoMonir/shamela_books_text_full.