An annotated corpus of Hebrew parliamentary proceedings containing over 35 million sentences from all the (plenary and committee) protocols held in the Israeli parliament
from 1992 to 2024.Sentences are annotated with various levels of linguistic information, including part-of-speech tags, morphological features, dependency⦠See the full description on the dataset page:
https://huggingface.co/datasets/HaifaCLGroup/KnessetCorpus.