gemma_3_12B_it_fp8_e4m3fn.safetensors - The fp8 converted text encoder from comfy, goes in
CLIP folder
gemma_3_12B_it_nvfp4_uncalibrated.safetensors - The nvfp4 converted (using comfyui kitchen utils) text encoder from comfy, goes in
CLIP folder.
NOTE: You will get warnings in the console, these can be ignored and will not affect your generation.
ltx-2-19b-dev-fp4_projections_only.safetensors - Extracted projections from LTX-2 model to allow loading with DualClipLoader node, goes in
CLIP folder
ltx-2-19b-dev-fp4_video_vae.safetensors - The video vae, can be loaded with
VaeLoader node, goes in
VAE folder
ltx-2-19b-dev-fp4_vocoder.safetensors - The vocoder model, not useful separately currently
For audio vae, use Kijai's audio vae upload
LTX2_audio_vae_bf16.safetensors, the one in this repo only contains the vocoder, but the audio vae is also needed. There was an oversight due to ComfyUI caching when it shouldn't.
When using a ComfyUI workflow which uses the original fp16 gemma 3 12b it model, simply select the text encoder from here instead.