Views
No views yet
nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-BF16 (hybrid Mamba2 + Attention + MoE, 120B total / ~12B active, 256k context)bias.md, explainability.md, custom modeling code) ships alongside the weights, as it did in the source repo. Sampler guidance from the main card applies unchanged: run it hot.