ScalaClassification
An MTEB dataset
Massive Text Embedding Benchmark
ScaLa a linguistic acceptability dataset for the mainland Scandinavian languages automatically constructed from dependency annotations in Universal Dependencies Treebanks.
Published as part of 'ScandEval: A Benchmark for Scandinavian Natural Language Processing'
Task category
t2c
Domains
Fiction, News, Non-fiction, Blog, Spoken, Web, Written