Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
SelectiveDPO-Llama3-8B-SFT-UFBinarized – AI Model by glorgao | AlphaNeural AI
You can deploy this model and start earning money today!
glorgao
/
SelectiveDPO-Llama3-8B-SFT-UFBinarized
like
0
transformers
safetensors
llama
text-generation
conversational
HuggingFaceH4/ultrafeedback_binarized
2502.09650
princeton-nlp/Llama-3-Base-8B-SFT
finetune
autotrain_compatible
text-generation-inference
endpoints_compatible
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
This model is fine-tuned from the princeton-nlp/Llama-3-Base-8B-SFT model using the
SelectiveDPO
on the Ultrafeedback_binarized dataset.
For the recipe to reproduce this model, please visit our
GitHub page
.