Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
MiMo-VL-7B-RL-2508-bnb-4bit-fp4 – AI Model by NangWeiLun | AlphaNeural AI
You can deploy this model and start earning money today!
NangWeiLun
/
MiMo-VL-7B-RL-2508-bnb-4bit-fp4
like
0
transformers
safetensors
qwen2_5_vl
image-to-text
quantization
4bit
bitsandbytes
bnb
memory-efficient
image-text-to-text
conversational
XiaomiMiMo/MiMo-VL-7B-RL-2508
quantized
mit
endpoints_compatible
4-bit
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
MiMo-VL-7B-RL-2508 — 4-bit BitsAndBytes Quantized
This is a
4-bit quantized
version of
XiaomiMiMo/MiMo-VL-7B-RL-2508
,
using the
BitsAndBytes
library.
Quantization reduces memory usage and makes it possible to run this model on consumer GPUs
(≤ 12 GB VRAM), at the cost of a small reduction in generation quality.
Quantization Details
Method
: BitsAndBytes (bnb)
Precision
: 4-bit (
fp4
)
Compute dtype
: bfloat16
Double quantization
: disabled
Format
:
safetensors