Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
Lyra4-Gutenberg-12B-4bpw – AI Model by Jellon | AlphaNeural AI
You can deploy this model and start earning money today!
Jellon
/
Lyra4-Gutenberg-12B-4bpw
like
0
transformers
safetensors
mistral
text-generation
conversational
jondurbin/gutenberg-dpo-v0.1
Sao10K/MN-12B-Lyra-v4
quantized
cc-by-nc-4.0
model-index
autotrain_compatible
text-generation-inference
endpoints_compatible
4-bit
exl2
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
4bpw exl2 quant of:
https://huggingface.co/nbeerbower/Lyra4-Gutenberg-12B
Lyra4-Gutenberg-12B
Sao10K/MN-12B-Lyra-v4
finetuned on
jondurbin/gutenberg-dpo-v0.1
.
Method
ORPO Finetuned using an RTX 3090 + 4060 Ti for 3 epochs.
Fine-tune Llama 3 with ORPO
Open LLM Leaderboard Evaluation Results
Detailed results can be found
here
Metric
Value
Avg.
19.63
IFEval (0-Shot)
22.12
BBH (3-Shot)
34.24
MATH Lvl 5 (4-Shot)
11.71
GPQA (0-shot)
9.17
MuSR (0-shot)
11.97
MMLU-PRO (5-shot)
28.57