Recent advances in LLMs have driven remarkable progress, yet their performance remains inconsistent on low-resource languages, highlighting challenges in equitable AI development.
This dataset demonstrates that while LLMs improve in Chuvash language understanding, fact-based knowledge about Chuvash literature remains a significant unresolved challenge.
The dataset is comprised of 100 questions designed to assess factual knowledge of Chuvash literature. The objective… See the full description on the dataset page:
https://huggingface.co/datasets/alexantonov/chuvash_llm_testset.