π Paper: Hinglish-Bench β Reference-Free Benchmark for LLM Hinglish Text
Generation
(gist preprint)
A reference-free benchmark for measuring how well LLMs generate natural
Roman-script Hinglish β the Hindi-English code-mixing that hundreds of
millions of Indians actually speak, type, and read online.
Reference-free by design. Hinglish has no canonical spelling and no single
"correct" rendering, so there are no gold references and no BLEU. Quality is⦠See the full description on the dataset page:
https://huggingface.co/datasets/saidutta69/hinglish-bench.