Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
Qwen2.5-VL-7B-Instruct-W4A16-G128 – AI Model by jeffcookio | AlphaNeural AI
You can deploy this model and start earning money today!
jeffcookio
/
Qwen2.5-VL-7B-Instruct-W4A16-G128
like
0
safetensors
qwen2_5_vl
compressed-tensors
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
Just a quick quantization of Qwen2.5-VL-7B-Instruct using
llm-compressor
. Used the example script with a
MAX_SEQUENCE
of 32768, truncation disabled (hit a bug in VLLM with this in the tokenizer), and
NUM_CALIBRATION_SAMPLES
of 512.