Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
DA-DPO_llava_v1.5_13B – AI Model by Artanic30 | AlphaNeural AI
You can deploy this model and start earning money today!
Artanic30
/
DA-DPO_llava_v1.5_13B
like
0
reinforcement-learning
en
2601.00623
liuhaotian/llava-v1.5-13b
finetune
apache-2.0
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
This is the official checkpoint released for the TMLR 2025 paper DA-DPO.
The training dataset can be found in
BPO
.
For model usage, please follow the instructions in
LLaVA-v1.5-13B
References
Model Paper