Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
qwama-0.5b-hh-rlhf-sft-chosen-trl-v4 – AI Model by lblaoke | AlphaNeural AI
You can deploy this model and start earning money today!
lblaoke
/
qwama-0.5b-hh-rlhf-sft-chosen-trl-v4
like
0
safetensors
qwen2
Dahoas/full-hh-rlhf
turboderp/Qwama-0.5B-Instruct
finetune
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
num_train_epochs: 1
learning_rate: 2e-4
total_batch_size: 32