Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
Lyra-Gutenberg-12b-exl3-4bpw – AI Model by Jellon | AlphaNeural AI
You can deploy this model and start earning money today!
Jellon
/
Lyra-Gutenberg-12b-exl3-4bpw
like
0
transformers
safetensors
mistral
text-generation
jondurbin/gutenberg-dpo-v0.1
nbeerbower/Lyra-Gutenberg-mistral-nemo-12B
quantized
cc-by-nc-4.0
model-index
autotrain_compatible
text-generation-inference
endpoints_compatible
4-bit
exl3
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
4bpw exl3 quant of:
https://huggingface.co/nbeerbower/Lyra-Gutenberg-mistral-nemo-12B
Lyra-Gutenberg-12B
Sao10K/MN-12B-Lyra-v1
finetuned on
jondurbin/gutenberg-dpo-v0.1
.
Method
Finetuned using an A100 on Google Colab for 3 epochs.
Fine-tune Llama 3 with ORPO
Open LLM Leaderboard Evaluation Results
Detailed results can be found
here
Metric
Value
Avg.
22.57
IFEval (0-Shot)
34.95
BBH (3-Shot)
36.99
MATH Lvl 5 (4-Shot)
8.31
GPQA (0-shot)
11.19
MuSR (0-shot)
14.76
MMLU-PRO (5-shot)
29.20