Nli-For-Simcse Marathi Dataset: High-Quality Marathi NLP Corpus
📌 Overview
The Nli-For-Simcse Marathi dataset is a meticulously curated collection of 274951 rows of Marathi text, ensuring linguistic accuracy and natural flow. Every sentence has been verified by native Marathi speakers to maintain contextual integrity and correctness.
This dataset is designed for semantic search, text classification, and various NLP tasks, making it a valuable resource for machine… See the full description on the dataset page: https://huggingface.co/datasets/Singhchandann/nli-for-simcse_marathi.