Search 3.2M models and datasets…
⌘K
Chat
Models
Datasets
Deploy
Pricing
Docs
Chat
Models
Datasets
Deploy
More
sprintduplicatequestions-pairclassification – Dataset by mteb | AlphaNeural AI
Is this your dataset? Claim it with the Hugging Face account that owns it.
mteb
/
sprintduplicatequestions-pairclassification
like
0
text-classification
semantic-similarity-classification
derived
monolingual
eng
unknown
n<1K
text
2502.13595
2210.07316
us
mteb
text
Views
No views yet
Dataset card
Files and Versions
Community
Use
Use this dataset
SprintDuplicateQuestions An MTEB dataset Massive Text Embedding Benchmark
Duplicate questions from the Sprint community.
Task category t2t
Domains Programming, Written
Reference
https://www.aclweb.org/anthology/D18-1131/
How to evaluate on this task
You can evaluate an embedding model on this dataset using the following code: import mteb
task = mteb.get_tasks(["SprintDuplicateQuestions"]) evaluator = mteb.MTEB(task)
model = mteb.get_model(YOUR_MODEL)… See the full description on the dataset page:
https://huggingface.co/datasets/mteb/sprintduplicatequestions-pairclassification
.