Multi-hop QA splits derived from MuSiQue and 2WikiMQA (as bundled in the
official CacheBlend repo), augmented
with extra distractor chunks per query so that each example carries 20
context passages instead of the original 10.
Designed to amplify the contrast between full_reuse (no cross-chunk
attention -> quality drop) and cacheblend (selective KV recompute ->
quality recovers) while preserving the multi-hop questions and gold
short answers.… See the full description on the dataset page:
https://huggingface.co/datasets/nicemyeong/cacheblend-rag-extended.