Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
Qwen2.5-0.5B-Math-GRPO – AI Model by tengfeima-ai | AlphaNeural AI
You can deploy this model and start earning money today!
tengfeima-ai
/
Qwen2.5-0.5B-Math-GRPO
like
0
safetensors
qwen2
math
reasoning
reinforcement-learning
grpo
rl
qwen2.5
text-generation
conversational
en
zwhe99/DeepMath-103K
2501.12948
tengfeima-ai/Qwen2.5-0.5B-Math-SFT
finetune
apache-2.0
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
No description or README available