"Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters"
(Snell, Lee, Xu, Kumar — 2024) — arxiv:2408.03314
The paper studies how to optimally scale inference-time computation in LLMs. The key finding: using a compute-optimal test-time strategy can improve efficiency by 4× compared… See the full description on the dataset page:
https://huggingface.co/datasets/ramu3405/math500-bon-prm-replication.