Parallel data for training a romanized-Hindi -> Devanagari transliteration
model that tolerates the messy, inconsistent way people actually type Hindi on
phones. Built to power a smart Hindi input method (IME), analogous to Chinese
smart-pinyin engines.
Each row is one (roman, devanagari) pair: roman is a plausible sloppy
human romanization, devanagari is the correct standard Hindi.