Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
Qwen.Qwen3-Coder-480B-A35B-Instruct-GGUF – AI Model by DevQuasar | AlphaNeural AI
You can deploy this model and start earning money today!
DevQuasar
/
Qwen.Qwen3-Coder-480B-A35B-Instruct-GGUF
like
0
quantized
Qwen/Qwen3-Coder-480B-A35B-Instruct
conversational
endpoints_compatible
gguf
template
us
text-generation
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
Q2 performace result on dual RTX5090 with 20 layers offloaded to GPU:
model
size
params
backend
ngl
test
t/s
qwen3moe ?B Q2_K - Medium
162.66 GiB
480.15 B
CUDA
20
pp512
90.09 ± 1.14
qwen3moe ?B Q2_K - Medium
162.66 GiB
480.15 B
CUDA
20
tg128
12.51 ± 0.11
'Make knowledge free for everyone'
Quantized version of:
Qwen/Qwen3-Coder-480B-A35B-Instruct
Buy Me a Coffee at ko-fi.com