As of June 2026, MTEB uses the training split of MKQA and SIB200 for evaluation. We recommend removing these two datasets from the training data.
@misc{f2llm-v2,
title={F2LLM-v2: Inclusive, Performant, and Efficient Embeddings for a Multilingual World},
author={Ziyin Zhang and Zihan Liao and Hang Yu and Peng Di and Rui Wang},
year={2026},
eprint={2603.19223},
archivePrefix={arXiv}… See the full description on the dataset page:
https://huggingface.co/datasets/codefuse-ai/F2LLM-v2.