Views
No views yet
| File | What it does |
|---|---|
MiniMax_H3_Turbo_v4_8step_T2VA_BennyDaBall.json | Text → video+audio with the Turbo v4 LoRA — 8 steps instead of 20, ~1.9× faster, audio stays clean |
| Setting | Value |
|---|---|
| LoRA | minimax_h3_turbo_v4_step600_ema_pruned_comfyui @ 1.0 (built-in LoraLoaderModelOnly) |
| Steps / sampler / scheduler | 8 · euler · beta (no sigma-shift node needed) |
| Guidance | CFG-free (BasicGuider) — write a Negative: line inside the prompt itself |
| Resolution | 1344×768 @ 24 fps |
| Length | frames on a 17k+5 grid: 124 ≈ 5 s, 243 ≈ 10 s, 362 ≈ 15 s (model max) |
.json onto the ComfyUI canvas.lora key not loaded warnings.| File | Size | Folder |
|---|---|---|
minimax_h3_fl2va_pruned_int8_convrot.safetensors | 19.5 GB | models/diffusion_models |
qwen3vl_32b_minimax_h3_nvfp4_awq.safetensors | 14.6 GB | models/text_encoders |
minimax_h3_video_vae_fp16.safetensors | 1.4 GB | models/vae |
minimax_h3_audio_vae_fp32.safetensors | 0.4 GB | models/vae |
minimax_h3_turbo_v4_step600_ema_pruned_comfyui.safetensors | 620 MB | models/loras |
0.0-0.6s: ...) → camera lock line → an audio timeline starting with room tone at 0.0s → Negative: line. Dialogue inside <d>[English] ... </d> is spoken verbatim with lip-sync. Give the first spoken line ≥ 1.2 s of lead-in and anchor t=0 with room tone. The demo prompt in the workflow shows the full pattern.