Views
No views yet
llama.cpp, LM Studio, and other GGUF-compatible runtimes.| File | Notes | Rough VRAM target |
|---|---|---|
nemotron-elastic-12b-Q4_K_S.gguf | Smaller 4-bit quant | ~10GB VRAM |
nemotron-elastic-12b-Q4_K_M.gguf | Better 4-bit quant | ~10GB+ VRAM |
HackerTwins/NVIDIA-Nemotron-Labs-3-Elastic-23B-A2.8B-GGUF