NanoBEIR translated in French (Microsoft API used).While in this collection we translate the datasets preserving the original formatting of zeta-alpha-ai, here the format was designed with Tom Aarsen to be compatible with the Sentence Transformers library. This dataset contains the 13 datasets composing NanoBEIR.It should also be noted that we have added a BM25 split (calculated using the bm25s library, using the same methodology as Tom Aarsen for the English dataset) so… See the full description on the dataset page:
https://huggingface.co/datasets/tomaarsen/NanoBEIR-fr-copy.