This dataset is a lightweight, evaluation-ready reformatting of the HLE-Verified benchmark created by the Skylenage Team.
Original work: Weiqi Zhai et al., "HLE-Verified: A Systematic Verification and Structured Revision of Humanity's Last Exam" (arXiv:2602.13964)
Original dataset: skylenage/HLE-Verified
Original repository: SKYLENAGE-AI/HLE-Verified
Converted from skylenage/HLE-Verified snapshot… See the full description on the dataset page:
https://huggingface.co/datasets/lmms-lab/HLE-Verified.