This dataset consists of parallel Hindi-English text pairs intended for training and fine-tuning machine translation models. The data has been collected from various sources to ensure diversity in sentence structures and vocabulary.
π Dataset Structure
Source Language: Hindi (hi)
Target Language: English (en)
Format: CSV / JSON / TXT
Columns:
source_text: The original Hindi sentence
target_text: The⦠See the full description on the dataset page: https://huggingface.co/datasets/zainabfatima097/My_Dataset.