Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
llama-8b-grpo-gsm8k – AI Model by ieayvaz | AlphaNeural AI
You can deploy this model and start earning money today!
ieayvaz
/
llama-8b-grpo-gsm8k
like
0
transformers
gguf
llama
text-generation-inference
unsloth
en
apache-2.0
endpoints_compatible
us
conversational
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
Uploaded model
Finetuned on GSM8K with GRPO. Quantization and Lora techniques are used for reducing VRAM usage.
Developed by:
ieayvaz
License:
apache-2.0
Finetuned from model :
unsloth/meta-llama-3.1-8b-instruct-bnb-4bit