This dataset contains sentences derived from medical case report abstracts
curated for adverse events. Split data and CoNLL formatting allows for the
training of language models, for named entity recognition. The dataset
includes entity annotations or labels. This subsect is the validation split.
The creation of the original PHEE dataset is detailed at: