Full Paper accepted at 1st Workshop on BHASHA: Benchmarks, Harmonization, Annotation, and Standardization for Human-Centric AI in Indian Languages at IJCNLP-AACL 2025.
Large language models (LLMs) are increasingly deployed in multilingual applications
but often generate plausible yet incorrect or misleading outputs, known as hallucinations.
While hallucination detection… See the full description on the dataset page:
https://huggingface.co/datasets/sambhashana/BHRAM-IL.