Views
No views yet
| branch | step | use |
|---|---|---|
main = step-1275 | last (1275/1275) | latest |
step-1200 | best eval_loss (0.317) | recommended |
step-1100 | 1100 | |
step-300 | 300 | early (mode-collapse) |
step-200 | 200 | |
step-100 | 100 | first working ckpt |
1from transformers import AutoModelForImageTextToText, AutoProcessor
2
3model = AutoModelForImageTextToText.from_pretrained(
4 "zhiyuanhucs/qwen3.5-4b-bc-sft-delta-force",
5 revision="step-1200", # change to any branch above
6 trust_remote_code=True,
7 torch_dtype="bfloat16",
8 device_map="auto",
9)
10processor = AutoProcessor.from_pretrained("zhiyuanhucs/qwen3.5-4b-bc-sft-delta-force", revision="step-1200", trust_remote_code=True)<|action_start|>X Y Z ; k1 ; k2 ; k3 ; k4 ; k5 ; k6<|action_end|>X Y Z = cumulative mouse delta + wheel, quantized at 5-pixel step; (one per 33-ms sub-frame)W/A/D/Shift/Ctrl/LB/RB/MB/Q/F/...)<|action_start|>, <|action_end|>, <|thought_start|>, <|thought_end|>scripts/42_check_format.py
in the training repo):| Metric | step-1200 |
|---|---|
contains <|action_start|> | 100% |
contains <|action_end|> | 100% |
| well-formed pair | 100% |
parses X Y Z as 3 ints | 100% |
uses ; separator | 100% |
exact_match (body, after strip <think></think>) | 14% |
| avg edit distance | 6.2 |