Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
distilGPT-NepSA – AI Model by raygx | AlphaNeural AI
You can deploy this model and start earning money today!
raygx
/
distilGPT-NepSA
like
0
transformers
tf
gpt2
text-classification
generated_from_keras_callback
apache-2.0
autotrain_compatible
endpoints_compatible
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
distilGPT-NepSA
This model is a fine-tuned version of
raygx/distilGPT-Nepali
on an unknown dataset. It achieves the following results on the evaluation set:
Train Loss: 0.6068
Validation Loss: 0.6592
Epoch: 1
Model description
More information needed
Intended uses & limitations
More information needed
Training and evaluation data
More information needed
Training procedure
Training hyperparameters
The following hyperparameters were used during training:
optimizer: {'name': 'AdamWeightDecay', 'learning_rate': 2e-05, 'decay': 0.0, 'beta_1': 0.9, 'beta_2': 0.999, 'epsilon': 1e-07, 'amsgrad': False, 'weight_decay_rate': 0.04}
training_precision: float32
Training results
Train Loss
Validation Loss
Epoch
0.8415
0.7254
0
0.6068
0.6592
1
Framework versions
Transformers 4.28.1
TensorFlow 2.11.0
Datasets 2.1.0
Tokenizers 0.13.3