Resource files for the mlnorm
multilingual lexical normalization toolkit. Not included in the PyPI
package due to size.
multilexnorm++/
~13 MB
MultiLexNorm 2026 benchmark data in .norm format (17 languages)
hunspell/
~42 MB
Hunspell spell-checker dictionaries (14… See the full description on the dataset page:
https://huggingface.co/datasets/hadung1802/mlnorm-resources.