Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
yuchenj_-_gpt2_774M_100B_FinewebEdu_hf-8bits – AI Model by RichardErkhov | AlphaNeural AI
You can deploy this model and start earning money today!
RichardErkhov
/
yuchenj_-_gpt2_774M_100B_FinewebEdu_hf-8bits
like
0
safetensors
gpt2
8-bit
bitsandbytes
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
Quantization made by Richard Erkhov.
Github
Discord
Request more models
gpt2_774M_100B_FinewebEdu_hf - bnb 8bits
Model creator:
https://huggingface.co/yuchenj/
Original model:
https://huggingface.co/yuchenj/gpt2_774M_100B_FinewebEdu_hf/
Original model description:
library_name: transformers datasets:
HuggingFaceFW/fineweb-edu
This is a GPT-2 (774M) model trained in llm.c for 100B tokens with cosine LR on Fineweb-Edu.
A lot more detailed info and observations are here:
https://x.com/Yuchenj_UW/status/1814703583453192272