This dataset is a community-driven collection of the Mon language (ISO 639-3: mon). It contains 15,000+ sentences and corresponding voice recordings collected via a Mon keyboard application. The goal is to provide high-quality open-source data to support Mon language integration into global AI systems like Google Translate, OpenAI Whisper, and ChatGPT.