Views
No views yet
model_world_size_1_rank_0.pt: model-only PyTorch state dict saved from a single-GPU FSDP/NO_SHARD run.extra_state_world_size_1_rank_0.pt: lightweight training state.huggingface/: tokenizer, processor, config, and generation config copied from the training runtime.save_pretrained() Hugging Face model directory. It is intended as a reproducibility checkpoint for the accompanying Agent-ChartQA training code.reward/overall: mean 0.9204, last-10-step mean 0.9371.reward/accuracy: mean 0.8553.reward/tool: 1.0 throughout the run.reward/format: 1.0 throughout the run.response_length/clip_ratio: 0.0 throughout the run.