Clustering of document titles from Allo Prof dataset. Clustering of 10 sets on the document topic.
You can evaluate an embedding model on this dataset using the following code:
import mteb
task = mteb.get_tasks(["AlloProfClusteringS2S.v2"])… See the full description on the dataset page:
https://huggingface.co/datasets/mteb/AlloProfClusteringS2S.v2.