The ESCWA-CS Corpus was collected over two days of meetings of the United Nations Economic and Social Commission for Western Asia (ESCWA) held in 2019.It contains intra-sentential code-switching between Arabic and English, with some speakers—particularly from Algeria, Tunisia, and Morocco—alternating between Arabic and French.
The dataset spans approximately 2.8 hours of speech, featuring dialectal Arabic and a Code Mixing Index… See the full description on the dataset page:
https://huggingface.co/datasets/MohamedRashad/simple-escwa.