Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
qwen-2.5-0.5b-grpo-rlcot-gsm8k – AI Model by nimishbongale | AlphaNeural AI
You can deploy this model and start earning money today!
nimishbongale
/
qwen-2.5-0.5b-grpo-rlcot-gsm8k
like
0
transformers
safetensors
qwen2
text-generation
conversational
en
openai/gsm8k
unsloth/Qwen2.5-0.5B-Instruct
finetune
apache-2.0
autotrain_compatible
text-generation-inference
endpoints_compatible
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
Trained on 100 epochs, has potential to scale upto 47-48% on GSM8k for a full run of 400 epochs.
image/png
image/png