Prepared contexts for fine-tuning and evaluating contextual Belarusian homograph stress
resolution. Each row describes one exact target occurrence and preserves its source text,
half-open character span, dictionary identifiers, provenance, quality tier, orthography,
grouped split, and sampling weight.
This is the training dataset for
fosters/homograph-bel-xlm-roberta-base.