This dataset is a Nano-style retrieval dataset for HAKARI-bench.
NanoRuMTEB is a compact Russian retrieval benchmark assembled from ruMTEB retrieval tasks. It includes Russian MIRACL, RIA News retrieval, and RuBQ retrieval; RuSciBench retrieval tasks are grouped separately under NanoMTEB-Misc.
queries = load_dataset(dataset_id, "queries", split=split)… See the full description on the dataset page:
https://huggingface.co/datasets/hakari-bench/NanoRuMTEB.