Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
llama-161M-100B-GGUF – AI Model by QuantFactory | AlphaNeural AI
You can deploy this model and start earning money today!
QuantFactory
/
llama-161M-100B-GGUF
like
0
abacaj/llama-161M-100B
quantized
endpoints_compatible
gguf
template
apache-2.0
us
text-generation
transformers
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
QuantFactory/llama-161M-100B-GGUF
This is quantized version of
abacaj/llama-161M-100B
created using llama.cpp
Model Description
Trained on 100B tokens.
1e-3 LR
0.1 wd
WSD scheduler with 10% decay
80% code, 10% NL, 10% instruction data
Dataset decontaminated against popular benchmarks following
bigcode
8x3090s 110~ hours
This is a
base
pretrained model and requires further fine tuning to be useful.
Model Details
openai/openai_humaneval
(greedy)
mbpp
(greedy)
9.2%
9.8%