Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
Qwythos-9B-v2-MLX-oQ8-mtp – AI Model by brainworkup | AlphaNeural AI
You can deploy this model and start earning money today!
brainworkup
/
Qwythos-9B-v2-MLX-oQ8-mtp
like
0
mlx
safetensors
qwen3_5
oq
quantized
omlx
oQ8
empero-ai/Qwythos-9B-v2
quantized
8-bit
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
Qwythos-9B-v2-MLX-oQ8-mtp
This model was quantized using
oQ
(oMLX v0.5.4.dev1) mixed-precision quantization.
Quantization details
Model type
: qwen3_5
Bits
: 8
Group size
: 64
Format
: MLX safetensors
Testing
I used these params recently for a complex reasoning task requring RAG and high-level thinking. Results were slow but exceptionally strong.
Added kwargs forced reasoning effort = Max
ACTIVE MODEL
Qwythos-9B-v2-MLX-oQ8-mtp
TEMPERATURE
0.6
MAX TOKENS
16384
MIN P
0.95 TOP K 20
REP. PENALTY
1.05
PRESENCE PENALTY
Default
THINKING
On (Unlimited)