Languages: English (en), Sanskrit (sa), Hindi (hi)
Dataset size: ~10K < n < ~100K records (chunks, verses, chapters, and articles combined)
This corpus is a comprehensive, highly structured dataset comprising the four primary pillars of ancient Indian literature (the Rig Veda, Sama Veda, Yajur Veda, and⦠See the full description on the dataset page:
https://huggingface.co/datasets/shinigamiRaj/IndianVedasOriginal.