BoBench — Tibetan Grammar, Literary & Historical Benchmark
Language: Classical & modern Tibetan (bo)Task: Multiple-choice question answering (MCQ) with domain-stratified reportingPrimary use: Intrinsic evaluation of Tibetan LLMs against SME-curated “ground truth” itemsMaintainer: Monlam AI
This card describes BoBench, a subject-matter-expert (SME)–governed benchmark for Tibetan grammar (བརྡ་སྤྲོད།), literature / poetics (རྩོམ་རིག · སྙན་ངག) and history (ལོ་རྒྱུས།), plus… See the full description on the dataset page:
https://huggingface.co/datasets/MonlamAI/Bo-bench-v1.0.0.