A small teaser subset (2.5k sentence pairs) of Garo–English parallel data released for experimentation and pipeline demos. The full corpus (200k pairs) remains proprietary. This teaser is not intended for benchmarking or production training.
Language pair: English (en) → Garo (grt)
Columns: source… See the full description on the dataset page:
https://huggingface.co/datasets/MWirelabs/garo-english-parallel-corpus.