This dataset is part of ArmBench-TextEmbed (Armenian Embedding Model Benchmark).
We manually created 185 high-quality query-document pairs covering diverse domains (finance, law, education) and use cases (customer service, sales). The annotations were reviewed by 2 independent reviewers to ensure accuracy of query-passage pairs. This dataset serves as our gold standard for Armenian retrieval evaluation.
If you use this dataset in… See the full description on the dataset page:
https://huggingface.co/datasets/Metric-AI/retrieval_dataset_hye.