NeuCLIR2023RetrievalHardNegatives
An MTEB dataset
Massive Text Embedding Benchmark
The task involves identifying and retrieving the documents that are relevant to the queries. The hard negative version has been created by pooling the 250 top documents per query from BM25, e5-multilingual-large and e5-mistral-instruct.