This repository contains a limited sample subset of an organic Egyptian Arabic (Masri) dataset.
Unlike standard web-scraped corpora or synthetic datasets, this data is generated entirely from the daily, spontaneous text and voice translation queries of native speakers translating between different Arabic dialects, as well as between Arabic dialects and other languages through our active mobile application. Therefore, it… See the full description on the dataset page:
https://huggingface.co/datasets/ebubekr53/organic-egyptian-arabic-dialect-dataset.