Beta
Explore
Marketplace
Neural Labs
Playground
Wallet
Docs
OpusparcusPC – Dataset by mteb | AlphaNeural AI
You can deploy this model and start earning money today!
mteb
/
OpusparcusPC
like
0
text-classification
semantic-similarity-classification
human-annotated
multilingual
deu
eng
fin
fra
rus
swe
cc-by-nc-4.0
n<1K
parquet
text
datasets
pandas
polars
mlcroissant
1809.06142
2502.13595
Views
No views yet
Model card
Files and Versions
Community
API
OpusparcusPC An MTEB dataset Massive Text Embedding Benchmark
Opusparcus is a paraphrase corpus for six European language: German, English, Finnish, French, Russian, and Swedish. The paraphrases consist of subtitles from movies and TV shows.
Task category t2t
Domains Spoken, Spoken
Reference
https://gem-benchmark.com/data_cards/opusparcus
How to evaluate on this task
You can evaluate an embedding model on this dataset using the following code: import… See the full description on the dataset page:
https://huggingface.co/datasets/mteb/OpusparcusPC
.