Beta
Explore
Marketplace
Neural Labs
Playground
Wallet
Docs
deepseek_r1_finetuned – AI Model by gimmy256 | AlphaNeural AI
You can deploy this model and start earning money today!
gimmy256
/
deepseek_r1_finetuned
like
0
transformers
pytorch
llama
text-generation
text-generation-inference
unsloth
trl
sft
conversational
en
FreedomIntelligence/medical-o1-reasoning-SFT
deepseek-ai/DeepSeek-R1
finetune
apache-2.0
autotrain_compatible
endpoints_compatible
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
Uploaded model
Developed by:
gimmy256
License:
apache-2.0
Finetuned from model :
unsloth/deepseek-r1-distill-llama-8b-unsloth-bnb-4bit
This llama model was trained 2x faster with
Unsloth
and Huggingface's TRL library.