Views
No views yet
step_8000.pt, the final checkpoint and best online-validation checkpoint from the local NanoChat EN/IT GPT-2-small-like WSD short-fast-decay web/wiki run 20260605_fresh-gpt2small-lr3e4-bs6-wsd-shortfastdecay8k-final5e6-webwiki.step_8000.pt~1.92Bstep_400080003.882301174948.535776128configs/eval/20260521_pretrain_minimal_en_it_webwiki_step11000.yamlval_loss_mixed: 5.3930ppl_mixed: 219.8592val_loss_en: 4.9928ppl_en: 147.3508val_loss_it: 4.1405ppl_it: 62.8313loop_rate: 0.400repeated_4gram_rate: 0.750distinct_2: 0.4706cloze_en_contains: 0.00cloze_it_contains: 0.12step_4000 -> mixed=5.1440step_7000 -> mixed=5.3313step_5000 -> mixed=5.3651step_8000 -> mixed=5.3930step_6000 -> mixed=5.5364step_8000 won the run's internal online validationstep_4000 won the external repo-native benchmark used to rank comparable releasesstep_8000 is the cleaner final checkpoint on repetition/diversity surface metricsstep_4000 remains the checkpoint we promote as the benchmark winnerstep_4000, this final checkpoint is behaviorally cleaner on several surface metrics:loop_rate: 0.400 vs 0.725repeated_4gram_rate: 0.750 vs 0.900distinct_2: 0.4706 vs 0.4251language_consistency_en: 1.00 vs 0.95val_loss_mixed: 5.3930 vs 5.1440source_loss_books_en: 5.1537source_loss_books_it: 5.1258source_loss_code: 8.3286source_loss_web_en: 6.2020source_loss_web_it: 6.4544source_loss_wiki_en: 3.9960source_loss_wiki_it: 3.6270training_config.yamltokenizer.jsontokenizer_meta.jsonstep_8000.ptstep_8000.safetensorsbest_validation.jsonmetrics.jsonleval_metrics.jsonlprobe_generations.jsonleval_summary.jsoncomparison.jsonbenchmark_report.mdbenchmark_metrics.jsonbenchmark_scores.jsonbenchmark_source_losses.jsonstep_4000.