Bright retrieval dataset with long documents.
You can evaluate an embedding model on this dataset using the following code:
import mteb
model = mteb.get_model(YOUR_MODEL)… See the full description on the dataset page:
https://huggingface.co/datasets/mteb/BrightLongRetrieval.