ARC-Bench: An Open-Ended Autonomous-Research Benchmark Across Five Scientific Domains
The benchmark released with AutoResearchClaw: Self-Reinforcing Autonomous Research with Human-AI Collaboration.
ARC-Bench is a 55-topic, open-ended autonomous-research benchmark spanning
five scientific domains. Each topic is not a fixed-input/fixed-output task — it is a
research question plus a structured briefing. A research agent (or a human) must
take a topic from question →… See the full description on the dataset page:
https://huggingface.co/datasets/AIMING-Lab-UNC/ARC-Bench.