This dataset contains a processed, clean Markdown version of the expansive Jewish text library sourced from the Sefaria project. It is specifically preprocessed step-by-step for use in Retrieval-Augmented Generation (RAG) pipelines and offline local AI assistants like Jude.
This section provides a description of the dataset fields, and additional information about the dataset structure such as criteria used to create the splits, relationships… See the full description on the dataset page:
https://huggingface.co/datasets/RockyCo/jude-judaic-data.