15 hand-crafted RAG (Retrieval-Augmented Generation) eval queries with ground-truth document IDs and reference answers. A small, fast benchmark for sanity-checking your retriever + answerer in CI before you reach for the heavyweight benchmarks.
Includes a negative-control query (no relevant docs exist) so you can verify your system doesn't hallucinate when retrieval comes back empty.