Views
No views yet
aking11/hyebert (6-layer, ~66M params) for
punctuation of Eastern Armenian participle clauses, as 4-class token labeling.
From the CODASSCA 2026 paper Sequence Labeling for Low-Resource Syntax.0 O · 1 COMMA_AFTER · 2 BUTH_AFTER · 3 REMOVE_COMMA| Benchmark | macro-F1 |
|---|---|
| Shtemaran 292 (clean textbook) | 0.3260 |
artifacts/; the notebook is in training/.1from transformers import AutoTokenizer, AutoModelForTokenClassification
2tok = AutoTokenizer.from_pretrained("AlbertHakobyan/hyebert-armenian-participle-punct")
3model = AutoModelForTokenClassification.from_pretrained("AlbertHakobyan/hyebert-armenian-participle-punct")