This dataset is a Nano-style retrieval dataset for HAKARI-bench.
NanoBEIR-it is the Italian language-specific component of MNanoBEIR. It groups compact BEIR-derived retrieval tasks for efficient evaluation of document ranking in that language.
queries = load_dataset(dataset_id, "queries", split=split)
corpus = load_dataset(dataset_id, "corpus"… See the full description on the dataset page:
https://huggingface.co/datasets/hakari-bench/NanoBEIR-it.