This dataset is a remastered version of
Reubencf/PolyglotText
prepared using Adaption's Adaptive Data platform.
Multilingual Sentences (Adaption)
9,999 sentences across 123 languages. A broad multilingual subset of
PolyglotText — originally derived from the
Tatoeba project — with Adaption-sharpened
enhanced_prompt / enhanced_completion / reasoning_trace columns.
Each row carries a source-language sentence, translations, and the
Adaption-processed fields.
Dataset size… See the full description on the dataset page: https://huggingface.co/datasets/Reubencf/Adaption-multilingual-sentences.