Views
No views yet
Qwen/Qwen3-30B-A3B-Base and is intended for efficient inference of Mixture-of-Experts (MoE) large language models with significantly reduced memory footprint.Qwen/Qwen3-30B-A3B-Basebitsmoe demobitsmoe eval --config configs/qwen3moe/eval.yaml1@misc{zhao2026bitsmoe,
2 title={{BitsMoE}: Efficient Spectral Energy-Guided Bit Allocation for {MoE} {LLM} Quantization},
3 author={Jiayu Zhao and Zihan Teng and Minhao Fan and Tianrui Ma and Wentao Ren and Song Chen and Weichen Liu},
4 year={2026},
5 eprint={2606.00079},
6 archivePrefix={arXiv},
7 primaryClass={cs.LG},
8 url={https://arxiv.org/abs/2606.00079}
9}