Parallel and monolingual corpora - 2 vol.
Description
A collection of parallel and monolingual corpora intended for applications in natural language processing, machine translation, and other artificial intelligence technologies.
Languages: Lithuanian, Spanish, Ukrainian, Norwegian, Swedish, and Danish.
The resource consists of monolingual corpora and parallel corpora covering texts from general, information technology, and legal domains.
Data volumes:
• Spanish… See the full description on the dataset page: https://huggingface.co/datasets/VSSA-SDSA/LT_MT_newC.