Benchmark dataset for evaluating ai-text-outline on Tibetan Buddhist texts.
Documents: 124
Source: BDRC/OpenPecha catalog transcriptions with human-confirmed section breakpoints
Evaluation repo: OpenPecha/ai-text-outline-benchmark
Package: ai-text-outline on PyPI
breakpoints — character indices where a new section begins in content
titles — the section title at each… See the full description on the dataset page:
https://huggingface.co/datasets/openpecha/ai-text-outline-benchmark.