Introduce
OpenSML, a series of
Open SMa
Ll Language Models. These models arcitecture are built stricly will Apple's
MLX framework.
The pre-training dataset is a slice of OpenWebText dataset with approximately 2.3 billion tokens.
OpenSMLis shared to advance open research by granting access to cutting-edge language models. However, because it’s trained on publicly sourced data and released without safety warranties, it may produce content that is inaccurate, harmful, biased, or otherwise objectionable. Users and developers should therefore conduct rigorous safety evaluations and put in place filtering or other safeguards that suit their specific use cases.
1@misc{zebrowski2025opensml,
2 title={OpenSML: A Family of Small Language Models},
3 author={William Zebrowski},
4 year={2025},
5 howpublished={\url{https://github.com/wzebrowski/opensml}}
6}