This model fine-tuned model of raygx/distilBERT-Nepali, revision no.: b35360e0cffb71ae18aaf4ea00ff8369964243a2
(This is because training is done in batches of data due to limited resources available)
This model is trained on
raygx/Nepali-Extended-Text-Corpus dataset.
This dataset is a mixture of cc100 and
raygx/Nepali-Text-Corpus.
Thus this model is trained on 10 times more data than its previous self.
Another change is, the tokenizer is different. Hence, it is a totally different model.
Training is done by running one epoch at once on a batch of data.
Thus, training is done for total 6 rounds.
So, there were total of 3 batches and 2 epochs.