This dataset is the official dataset accompanying "Can LLMs Help Uncover Insights about LLMs? A Large-Scale, Evolving Literature Analysis of Frontier LLMs (ACL 2025 Main)".
It contains automatically extracted experimental results from arXiv papers along with various attributes focusing on frontier Language Models (LLMs).
By capturing model performance metrics, experimental conditions, and related attributes in a structured format, this dataset enables comprehensive… See the full description on the dataset page:
https://huggingface.co/datasets/jungsoopark/LLMEvalDB.