Token-level hallucination annotations on LLM answers grounded in prose
context, drawn from two public RAG hallucination resources and mapped into one
unified taxonomy. This is the prose counterpart to the structured-context
(code, tool output, documents)
collection — together they let a single detector be trained across modalities.
Two sources sit side by side, distinguished by the dataset field: