Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
qwen3.5-9b-my-style-dpo-unsloth – AI Model by miaota | AlphaNeural AI
You can deploy this model and start earning money today!
miaota
/
qwen3.5-9b-my-style-dpo-unsloth
like
0
safetensors
qwen3_5
qwen3.5
dpo
lora
unsloth
aesthetic-alignment
zh
Qwen/Qwen3.5-9B
adapter
apache-2.0
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
miaota/qwen3.5-9b-my-style-dpo-unsloth
基于 Qwen3.5-9B 的"审美映射"模型 DPO 阶段产物(
Unsloth 训练
)。 DPO 学习的是 subtle aesthetic preference——为什么这个表达比另一个更妙。