Views
No views yet
| Model | Layers | Hidden Size | Params | Max Seq Len | Learning Rate | Batch Size | Train Steps |
|---|---|---|---|---|---|---|---|
| Dialog-KoELECTRA-Small | 12 | 256 | 14M | 128 | 1e-4 | 512 | 700K |
| NSMC (acc) | Question Pair (acc) | Korean-Hate-Speech (F1) | Naver NER (F1) | KorNLI (acc) | KorSTS (spearman) | |
|---|---|---|---|---|---|---|
| DistilKoBERT | 88.60 | 92.48 | 60.72 | 84.65 | 72.00 | 72.59 |
| Dialog-KoELECTRA-Small | 90.01 | 94.99 | 68.26 | 85.51 | 78.54 | 78.96 |
| corpus name | size | |
|---|---|---|
| dialog | Aihub Korean dialog corpus | 7GB |
| NIKL Spoken corpus | ||
| Korean chatbot data | ||
| KcBERT | ||
| written | NIKL Newspaper corpus | 15GB |
| namuwikitext |
| vocabulary size | unused token size | limit alphabet | min frequency |
|---|---|---|---|
| 40,000 | 500 | 6,000 | 3 |