Views
No views yet
digits → 108年 / zh-num → 一百零八年) in the decoder-prefix channel.language Chinese format digits<asr_text> / … format zh-num<asr_text>.scripts/probe_numeral_flip.py):| Panel (audio) | pair flip | digits compliance | zh-num compliance |
|---|---|---|---|
| CommonVoice zh-TW | 0.73 | 0.77 | 0.93 |
| NTUML2021 | 0.27 | 0.43 | 0.70 |
| Benchmark | TEA-ASR-1.1-mini-fmt | TEA-ASR-1.1-mini | Qwen3-ASR-0.6B | Breeze-ASR-25 |
|---|---|---|---|---|
| CommonVoice 19 (zh-TW) | 5.27 | 5.12 | 5.79 | 8.03 |
| ASCEND (zh-en) | 11.25 | 11.20 | 12.54 | 17.53 |
| CSZS (zh-en) | 14.83 | 12.51 | 16.03 | 12.18 |
| NTUML2021 | 7.44 | 7.53 | 11.03 | 7.50 |
< 10 hours of public training audio, rank-16 decoder LoRA + low-LR encoder LoRA merged
into one drop-in checkpoint, Traditional output from the model's own tokenizer (no runtime
post-processing; decode verified bit-exact on 152k+ sequences). The fmt recipe adds
numeral-convention counterfactual pairs (same audio, both conventions, opposite tags) mined from
the same corpora at zero extra audio budget.1@misc{teaasr2026,
2 title = {Tokenizer-First Adaptation of Mandarin ASR to Taiwan Mandarin},
3 author = {TEA-ASR contributors},
4 year = {2026},
5 note = {TEA-ASR (Taiwan Everyday Audio); adapted from Qwen3-ASR}
6}