MILQA is a Hungarian machine reading comprehension, specifically, question answering (QA) benchmark database. In English, the most basic resource for the task is the Stanford Question Answering Dataset (SQuAD). The database was largely built following the principles of SQuAD 2.0, and is therefore characterized by the following:
Excerpts from high quality Wikipedia articles are used as context for the questions (free-to-use texts… See the full description on the dataset page:
https://huggingface.co/datasets/SzegedAI/MILQA.