Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
GPT-NepSA-T2 – AI Model by raygx | AlphaNeural AI
You can deploy this model and start earning money today!
raygx
/
GPT-NepSA-T2
like
0
transformers
tf
gpt2
text-classification
generated_from_keras_callback
raygx/Nepali-GPT2-CausalLM
finetune
autotrain_compatible
endpoints_compatible
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
GPT-NepSA-T2
This model is a fine-tuned version of
raygx/Nepali-GPT2-CausalLM
on an unknown dataset. It achieves the following results on the evaluation set:
Train Loss: 0.4349
Validation Loss: 0.7017
Epoch: 4
Model description
More information needed
Intended uses & limitations
More information needed
Training and evaluation data
More information needed
Training procedure
Training hyperparameters
The following hyperparameters were used during training:
optimizer: {'name': 'AdamWeightDecay', 'learning_rate': 5e-06, 'decay': 0.0, 'beta_1': 0.9, 'beta_2': 0.999, 'epsilon': 1e-07, 'amsgrad': False, 'weight_decay_rate': 0.005}
training_precision: float32
Training results
Train Loss
Validation Loss
Epoch
0.7875
0.6565
0
0.6364
0.6237
1
0.5681
0.6358
2
0.5018
0.6602
3
0.4349
0.7017
4
Framework versions
Transformers 4.31.0
TensorFlow 2.12.0
Datasets 2.13.1
Tokenizers 0.13.3