Per-step training metrics from a four-dimensional grid of small language-model pretraining runs, produced for the scaling-law study "Tokens-per-Parameter Coverage Is Critical for Robust LLM Scaling Law Extrapolation". The archive was moved here from the lab-v2/4d_llm_grid code repository to keep the large binary metric files out of Git.