Nepali Acts corpus — official legal documents of Nepal in Nepali language.
This dataset is a cleaned, chunked text corpus (approx. 400–500 words per chunk) constructed from official Nepali legal documents.It is designed for tokenizer training and for continued pre‑training of language models in the legal domain.
नेपाल सरकार, कानून, न्याय तथा संसदीय मामिला मन्त्रालय(Government of… See the full description on the dataset page:
https://huggingface.co/datasets/chhatramani/nepal_legal_acts_corpus_nepali.