Language: German Training data: German STS benchmark train and dev set Eval data: German STS benchmark test set Infrastructure: 1x V100 GPU Published: August 12th, 2021
Details
We trained a gbert-large model on the task of estimating semantic similarity of German-language text pairs. The dataset is a machine-translated version of the STS benchmark, which is available here.