Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
OrpoLlama-3-8B-AWQ – AI Model by solidrust | AlphaNeural AI
You can deploy this model and start earning money today!
solidrust
/
OrpoLlama-3-8B-AWQ
like
0
transformers
safetensors
llama
text-generation
4-bit
AWQ
autotrain_compatible
endpoints_compatible
orpo
llama 3
rlhf
sft
conversational
en
mlabonne/orpo-dpo-mix-40k
mlabonne/OrpoLlama-3-8B
quantized
other
text-generation-inference
awq
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
mlabonne/OrpoLlama-3-8B AWQ
Model creator:
mlabonne
Original model:
OrpoLlama-3-8B
Model Summary
This is an ORPO fine-tune of
meta-llama/Meta-Llama-3-8B
on 1k samples of
mlabonne/orpo-dpo-mix-40k
created for
this article
.
It's a successful fine-tune that follows the ChatML template!
Try the demo
:
https://huggingface.co/spaces/mlabonne/OrpoLlama-3-8B
🔎 Application
This model uses a context window of 8k. It was trained with the ChatML template.
🏆 Evaluation
Nous
OrpoLlama-4-8B outperforms Llama-3-8B-Instruct on the GPT4All and TruthfulQA datasets.