Beta
Explore
Marketplace
Neural Labs
Playground
Wallet
Docs
Mistral-Base-7B-DPO_clean – AI Model by ComparisonPO | AlphaNeural AI
You can deploy this model and start earning money today!
ComparisonPO
/
Mistral-Base-7B-DPO_clean
like
0
safetensors
mistral
trl-lib/ultrafeedback_binarized
alignment-handbook/zephyr-7b-sft-full
finetune
mit
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
DPO model excluding the noisy preference pairs for Mistral-Base under trl/ultradeedback_binarized finetuning.