A balanced 1000-prompt subsample of bench-llm/or-bench (or-bench-80k config), used for overrefusal evaluation in MR-Eval.
Construction
Source: bench-llm/or-bench, config or-bench-80k (80,359 prompts).
Sampled 100 prompts per category for each of the 10 categories (random seed 42).
Final order shuffled.