😊 Training Dense Retrievers with Multiple Positive Passages
If you like our work,
This repository contains the dataset used in our paper: "Training Dense Retrievers with Multiple Positive Passages".
Using the Qwen3-32B, we annotate utility on the MS MARCO dataset.
For NQ and hybrid test set, see Utility_focused_annotation
Our paper presents a systematic study of multi-positive optimization objectives for dense retrieval… See the full description on the dataset page:
https://huggingface.co/datasets/Wanglanhuajiaofen/MSMARCO-annotation.