Beta
Explore
Marketplace
Neural Labs
Playground
Wallet
Docs
Llama3-8B-RDPO – AI Model by RTO-RL | AlphaNeural AI
You can deploy this model and start earning money today!
RTO-RL
/
Llama3-8B-RDPO
like
0
safetensors
llama
HuggingFaceH4/ultrafeedback_binarized
OpenRLHF/Llama-3-8b-sft-mixture
finetune
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
Base model:
OpenRLHF/Llama-3-8b-sft-mixture
Preference dataset:
HuggingFaceH4/ultrafeedback_binarized