Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
One-Shot-RLVR-Qwen2.5-Math-1.5B-7.5k-MATH – AI Model by ypwang61 | AlphaNeural AI
You can deploy this model and start earning money today!
ypwang61
/
One-Shot-RLVR-Qwen2.5-Math-1.5B-7.5k-MATH
like
0
transformers
safetensors
qwen2
text-generation
conversational
ypwang61/One-Shot-RLVR-Datasets
2504.20571
Qwen/Qwen2.5-Math-1.5B
finetune
apache-2.0
autotrain_compatible
text-generation-inference
endpoints_compatible
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
This repository contains the model presented in
Reinforcement Learning for Reasoning in Large Language Models with One Training Example
.
Code:
https://github.com/ypwang61/One-Shot-RLVR