This dataset contains a teacher-verified Sinhala sentence–gloss parallel corpus developed as part of the SSLGen: Sinhala Sign Language Text-to-Gloss-to-Video Translation Framework.
The corpus was created to support research on Sinhala Sign Language (SSL) translation, particularly the task of converting written Sinhala sentences into Sinhala Sign Language gloss sequences.
The dataset contains 3,363 Sinhala sentence–gloss pairs collected from… See the full description on the dataset page:
https://huggingface.co/datasets/GeethmaYasashwi/SSL-Gen_Sinhala-Sentences.