Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
IndicCrosslingualSTS – Dataset by mteb | AlphaNeural AI
You can deploy this model and start earning money today!
mteb
/
IndicCrosslingualSTS
like
0
sentence-similarity
semantic-similarity-scoring
expert-annotated
multilingual
asm
ben
eng
guj
hin
kan
mal
mar
ory
pan
tam
tel
urd
cc0-1.0
1K<n<10K
parquet
Views
No views yet
Model card
Files and Versions
Community
API
IndicCrosslingualSTS An MTEB dataset Massive Text Embedding Benchmark
This is a Semantic Textual Similarity testset between English and 12 high-resource Indic languages.
Task category t2t
Domains News, Non-fiction, Web, Spoken, Government, Written, Spoken Reference
https://huggingface.co/datasets/jaygala24/indic_sts
How to evaluate on this task
You can evaluate an embedding model on this dataset using the following code: import mteb
task =… See the full description on the dataset page:
https://huggingface.co/datasets/mteb/IndicCrosslingualSTS
.