A computational resource for Purépecha (language isolate, ~128,344 speakers, Michoacán, Mexico) featuring 8,511 sentence pairs compiled from five sources with quantified orthographic and morphological variation, plus a reproducible linguistic annotation pipeline.
Status: v1.0.0-rc1 (pre-release candidate)License:… See the full description on the dataset page:
https://huggingface.co/datasets/CeciGonSer/purepecha-spanish-corpus.