Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
Qwen2.5-0.5B-Instruct-Align-Anything-DPO – AI Model by ll922 | AlphaNeural AI
You can deploy this model and start earning money today!
ll922
/
Qwen2.5-0.5B-Instruct-Align-Anything-DPO
like
0
safetensors
qwen2
PKU-Alignment/align-anything
Qwen/Qwen2.5-0.5B-Instruct
finetune
apache-2.0
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
DPO training is performed using the
Align-Anything
framework, with the
PKU-Alignment/align-anything
text-to-text dataset.
DPO training report:
https://api.wandb.ai/links/nlp-amct/uifw66p5