Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
gemma3-1b-reasoning-sft – AI Model by thiagobcunha | AlphaNeural AI
You can deploy this model and start earning money today!
thiagobcunha
/
gemma3-1b-reasoning-sft
like
0
transformers
safetensors
text-generation-inference
unsloth
gemma3_text
trl
en
unsloth/gemma-3-1b-it
finetune
apache-2.0
endpoints_compatible
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
Gemma 3 1B — Reasoning SFT (LoRA)
Gemma 3 1B model fine-tuned using
Supervised Fine-Tuning (SFT)
to generate responses with
structured reasoning
, mandatorily using the
<reasoning>
and
<answer>
tags.
Base:
google/gemma-3-1b-it
Finetuned:
thiagobcunha/gemma3-1b-reasoning-sft