Stsb Marathi Dataset: High-Quality Marathi NLP Corpus
๐ Overview
The Stsb Marathi dataset is a meticulously curated collection of 8628 rows of Marathi text, ensuring linguistic accuracy and natural flow. Every sentence has been verified by native Marathi speakers to maintain contextual integrity and correctness.
This dataset is designed for semantic search, text classification, and various NLP tasks, making it a valuable resource for machine learning models dealing withโฆ See the full description on the dataset page: https://huggingface.co/datasets/Singhchandann/stsb_marathi.