Views
No views yet
0.15| Tham số | Giá trị |
|---|---|
| Max sequence length | 256 |
| Batch size | 256 |
| Epochs | 10 |
| Learning rate | 0.0001 |
| MLM probability | 0.15 |
1from transformers import BertForPreTraining, BertTokenizerFast
2import torch
3
4tokenizer = BertTokenizerFast.from_pretrained("ducanhdinh/jepa_proof_bert")
5model = BertForPreTraining.from_pretrained("ducanhdinh/jepa_proof_bert")
6
7encoded = tokenizer("Hello world!", return_tensors="pt")
8with torch.no_grad():
9 output = model(**encoded)