Dataset Card for OPUS EUconst
Dataset Summary
Languages: 21
Bitexts: 210
Number of files: 986
Number of tokens: 3.01M
Sentence fragments: 0.22M
The underlying task is machine translation.
Czech (cs)
Danish (da)
German (de)
Greek (el)
English (en)
Spanish (es)
Estonian (et)
Finnish (fi)
French… See the full description on the dataset page:
https://huggingface.co/datasets/Helsinki-NLP/euconst.