A contextualized sentence-level (query2chunk) retrieval eval built from SEC 8-K
restructuring filings (EDGAR), for the
Chunk-level Retrieval Eval
collection.
Why this benchmark exists. Standard passage benchmarks (e.g. DAPR) don't discriminate
contextual embedding models, because their answer passages already contain the
distinguishing entity and have no near-duplicate distractors. This dataset is built to
stress… See the full description on the dataset page:
https://huggingface.co/datasets/bowang0911/EDGAR8KContextChunkRetrieval.