Views
No views yet
nvidia/Nemotron-Cascade-2-30B-A3B quantized pack, published as Nemotron-Cascade-2-30B-A3B-TurboQuant-MLX-6bit.pipeline_tag: text-generation.
This is a Mixture-of-Experts (MoE) model — a subset of experts is active per token; total and active parameter counts differ.nvidia/Nemotron-Cascade-2-30B-A3B; all credit for the original model, training, and weights belongs to the upstream authors. This repo republishes a quantized conversion of those weights only.