Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
nbeerbower_-_mistral-nemo-gutenberg-12B-v4-gguf – AI Model by RichardErkhov | AlphaNeural AI
You can deploy this model and start earning money today!
RichardErkhov
/
nbeerbower_-_mistral-nemo-gutenberg-12B-v4-gguf
like
0
conversational
endpoints_compatible
gguf
template
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
Quantization made by Richard Erkhov.
Github
Discord
Request more models
mistral-nemo-gutenberg-12B-v4 - GGUF
Model creator:
https://huggingface.co/nbeerbower/
Original model:
https://huggingface.co/nbeerbower/mistral-nemo-gutenberg-12B-v4/
Name
Quant method
Size
mistral-nemo-gutenberg-12B-v4.Q2_K.gguf
Q2_K
4.46GB
mistral-nemo-gutenberg-12B-v4.IQ3_XS.gguf
IQ3_XS
4.94GB
mistral-nemo-gutenberg-12B-v4.IQ3_S.gguf
IQ3_S
5.18GB
mistral-nemo-gutenberg-12B-v4.Q3_K_S.gguf
Q3_K_S
5.15GB
mistral-nemo-gutenberg-12B-v4.IQ3_M.gguf
IQ3_M
5.33GB
mistral-nemo-gutenberg-12B-v4.Q3_K.gguf
Q3_K
5.67GB
mistral-nemo-gutenberg-12B-v4.Q3_K_M.gguf
Q3_K_M
5.67GB
mistral-nemo-gutenberg-12B-v4.Q3_K_L.gguf
Q3_K_L
6.11GB
mistral-nemo-gutenberg-12B-v4.IQ4_XS.gguf
IQ4_XS
6.33GB
mistral-nemo-gutenberg-12B-v4.Q4_0.gguf
Q4_0
6.59GB
mistral-nemo-gutenberg-12B-v4.IQ4_NL.gguf
IQ4_NL
6.65GB
mistral-nemo-gutenberg-12B-v4.Q4_K_S.gguf
Q4_K_S
6.63GB
mistral-nemo-gutenberg-12B-v4.Q4_K.gguf
Q4_K
6.96GB
mistral-nemo-gutenberg-12B-v4.Q4_K_M.gguf
Q4_K_M
6.96GB
mistral-nemo-gutenberg-12B-v4.Q4_1.gguf
Q4_1
7.26GB
mistral-nemo-gutenberg-12B-v4.Q5_0.gguf
Q5_0
7.93GB
mistral-nemo-gutenberg-12B-v4.Q5_K_S.gguf
Q5_K_S
7.93GB
mistral-nemo-gutenberg-12B-v4.Q5_K.gguf
Q5_K
8.13GB
mistral-nemo-gutenberg-12B-v4.Q5_K_M.gguf
Q5_K_M
8.13GB
mistral-nemo-gutenberg-12B-v4.Q5_1.gguf
Q5_1
8.61GB
mistral-nemo-gutenberg-12B-v4.Q6_K.gguf
Q6_K
9.37GB
mistral-nemo-gutenberg-12B-v4.Q8_0.gguf
Q8_0
12.13GB
Original model description:
license: apache-2.0 library_name: transformers base_model:
TheDrummer/Rocinante-12B-v1 datasets:
jondurbin/gutenberg-dpo-v0.1 model-index:
name: mistral-nemo-gutenberg-12B-v4 results:
task: type: text-generation name: Text Generation dataset: name: IFEval (0-Shot) type: HuggingFaceH4/ifeval args: num_few_shot: 0 metrics:
type: inst_level_strict_acc and prompt_level_strict_acc value: 23.79 name: strict accuracy source: url:
https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=nbeerbower/mistral-nemo-gutenberg-12B-v4
name: Open LLM Leaderboard
task: type: text-generation name: Text Generation dataset: name: BBH (3-Shot) type: BBH args: num_few_shot: 3 metrics:
type: acc_norm value: 31.97 name: normalized accuracy source: url:
https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=nbeerbower/mistral-nemo-gutenberg-12B-v4
name: Open LLM Leaderboard
task: type: text-generation name: Text Generation dataset: name: MATH Lvl 5 (4-Shot) type: hendrycks/competition_math args: num_few_shot: 4 metrics:
type: exact_match value: 10.95 name: exact match source: url:
https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=nbeerbower/mistral-nemo-gutenberg-12B-v4
name: Open LLM Leaderboard
task: type: text-generation name: Text Generation dataset: name: GPQA (0-shot) type: Idavidrein/gpqa args: num_few_shot: 0 metrics:
type: acc_norm value: 8.84 name: acc_norm source: url:
https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=nbeerbower/mistral-nemo-gutenberg-12B-v4
name: Open LLM Leaderboard
task: type: text-generation name: Text Generation dataset: name: MuSR (0-shot) type: TAUR-Lab/MuSR args: num_few_shot: 0 metrics:
type: acc_norm value: 13.2 name: acc_norm source: url:
https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=nbeerbower/mistral-nemo-gutenberg-12B-v4
name: Open LLM Leaderboard
task: type: text-generation name: Text Generation dataset: name: MMLU-PRO (5-shot) type: TIGER-Lab/MMLU-Pro config: main split: test args: num_few_shot: 5 metrics:
type: acc value: 28.62 name: accuracy source: url:
https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=nbeerbower/mistral-nemo-gutenberg-12B-v4
name: Open LLM Leaderboard
mistral-nemo-gutenberg-12B-v4
TheDrummer/Rocinante-12B-v1
finetuned on
jondurbin/gutenberg-dpo-v0.1
.
Method
Finetuned using an A100 on Google Colab for 3 epochs.
Fine-tune Llama 3 with ORPO
Open LLM Leaderboard Evaluation Results
Detailed results can be found
here
Metric
Value
Avg.
19.56
IFEval (0-Shot)
23.79
BBH (3-Shot)
31.97
MATH Lvl 5 (4-Shot)
10.95
GPQA (0-shot)
8.84
MuSR (0-shot)
13.20
MMLU-PRO (5-shot)
28.62