Beta
Explore
Marketplace
Neural Labs
Playground
Wallet
Docs
enwik8 – Dataset by LTCB | AlphaNeural AI
You can deploy this model and start earning money today!
LTCB
/
enwik8
like
0
fill-mask
text-generation
language-modeling
masked-language-modeling
no-annotation
found
monolingual
original
en
mit
10K<n<100K
us
Views
No views yet
Model card
Files and Versions
Community
API
The dataset is based on the Hutter Prize (
http://prize.hutter1.net
) and contains the first 10^8 bytes of English Wikipedia in 2006 in XML