Views
No views yet
--model_version original) on Edit-2511 weights with surgical Tier B+ targeting (image-stream AdaLN modulation + image-stream MLP only). Designed to preserve native multi-reference editing and InstantX ControlNet compatibility.t01nstylet01nstyle_qie2511_t2i_surgical-000018.safetensors — global minimum loss + center of the sweet spot.t01nstyle_qie2511_t2i_surgical_final.safetensors corresponds to epoch 25.| epoch | loss | epoch | loss | |
|---|---|---|---|---|
| 1 | 0.07772 | 14 | 0.07651 | |
| 2 | 0.07733 | 15 | 0.07539 | |
| 3 | 0.07672 | 16 | 0.07353 | |
| 4 | 0.07606 | 17 | 0.07477 | |
| 5 | 0.07701 | 18 | 0.07351 (min) | |
| 6 | 0.07731 | 19 | 0.07443 | |
| 7 | 0.07579 | 20 | 0.07462 | |
| 8 | 0.07678 | 21 | 0.07602 | |
| 9 | 0.07644 | 22 | 0.07664 | |
| 10 | 0.07666 | 23 | 0.07661 | |
| 11 | 0.07833 | 24 | 0.07404 | |
| 12 | 0.07759 | 25 | 0.07389 | |
| 13 | 0.07379 |
controlnet_conditioning_scale 0.6-0.8)--model_version original on Edit-2511 weights--network_argsimg_mod.1, img_mlp.net.0.proj, img_mlp.net.2 (image stream only, attention untouched)to_q/k/v, to_out.0, add_q/k/v_proj, to_add_out) are bit-identical to base Edit-2511. Multi-reference image embeddings flow through joint attention unchanged. ControlNet residuals land on unchanged hidden states. The LoRA modulates only per-block style injection (AdaLN) and image-stream FFN output.t01nstyle_qie2511_t2i_surgical-000001.safetensors … -000024.safetensors — checkpoints per epoch (1–24)t01nstyle_qie2511_t2i_surgical_final.safetensors — epoch 25 finaltrain.sh — exact training commanddataset.toml — dataset configurationtraining.log — full training logtensorboard/ — tensorboard event files for loss curves