Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
Qwen3.5-0.8B-GRPO-Math – AI Model by zosmaai | AlphaNeural AI
You can deploy this model and start earning money today!
zosmaai
/
Qwen3.5-0.8B-GRPO-Math
like
0
safetensors
qwen3_5_text
reasoning
math
grpo
reinforcement-learning
rlvr
qwen3.5
text-generation
conversational
en
gsm8k
zosmaai/Qwen3.5-0.8B-GRPO-Math-Dataset
2402.03300
Qwen/Qwen3.5-0.8B
finetune
apache-2.0
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
No description or README available