Views
No views yet
track_encoder + patch-embed track slot),
head trainable, on a small 50-clip 720p synthetic set. Fixed track-ID sampling
(WANTRACK_FIXED_SAMPLE=1) so each clip presents an identical sparse conditioning pattern.lr 1e-4 (constant), 2000 steps, global batch 8, flow_shift 6, sparse conditioning
(WANTRACK_SPARSE=1, EXTRA_RANDOM=20), d64 track-ID embedding + bias track encoder.trackwan_14b_i2v_d64_bias_init (Wan2.1-I2V-14B-720P grafted to 52-ch TrackWan DiT).dcp/__<rank>_0.distcp (8 ranks, HSDP 2×4). It is not a from_pretrained
model — to load weights, consolidate via the FastVideo export step (03_export.sh).checkpoint-2000/ under a training output_dir and setting
resume_from_checkpoint: latest.data_pipeline/720_stage_1/ recipe.