IndicQA is a manually curated cloze-style reading comprehension dataset that can be used for evaluating question-answering models in 11 Indic languages. It is repurposed retrieving relevant context for each question.
You can evaluate an embedding model on this dataset using the… See the full description on the dataset page:
https://huggingface.co/datasets/mteb/IndicQARetrieval.