RoSTSC is a Romanian Semantic Textual Similarity (STS) dataset designed for evaluating and training sentence embedding models. It contains pairs of Romanian sentences along with similarity scores that indicate the degree of semantic equivalence between them.
sentence1: The first sentence in the pair.
sentence2: The second sentence in the pair.
score: A numerical value representing the semantic similarity between the… See the full description on the dataset page:
https://huggingface.co/datasets/BlackKakapo/RoSTSC.