Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
Qwen2.5-3B-Instruct-Reasoning-gsm8k-lora-v1 – AI Model by nomadicsynth | AlphaNeural AI
You can deploy this model and start earning money today!
nomadicsynth
/
Qwen2.5-3B-Instruct-Reasoning-gsm8k-lora-v1
like
0
peft
safetensors
text-generation-inference
unsloth
qwen2
trl
text-generation
conversational
en
openai/gsm8k
unsloth/Qwen2.5-3B-Instruct-unsloth-bnb-4bit
adapter
apache-2.0
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
Uploaded model
Developed by:
nomadicsynth
License:
apache-2.0
Finetuned from model:
unsloth/Qwen2.5-3B-Instruct-unsloth-bnb-4bit
Training Notebook:
Qwen2.5_(3B)-GRPO.ipynb
This qwen2 model was trained 2x faster with
Unsloth
and Huggingface's TRL library.