Views
No views yet
uitnlp/CafeBERT for binary Vietnamese clickbait detection.StratifiedGroupKFold with seed 42.True.| Metric | Mean ± sample std |
|---|---|
| Test Macro-F1 | 0.8047 ± 0.0060 |
| Test accuracy | 0.8236 ± 0.0074 |
| Dev Macro-F1 | 0.8256 ± 0.0053 |
| seed | dev_macro_f1 | test_macro_f1 | test_accuracy |
|---|---|---|---|
| 22 | 0.8312 | 0.8043 | 0.8246 |
| 42 | 0.8249 | 0.7989 | 0.8158 |
| 202 | 0.8207 | 0.8108 | 0.8304 |
0: non-clickbait1: clickbait1from transformers import AutoModelForSequenceClassification, AutoTokenizer
2
3model_id = "BaoNhan/cafebert-ViClickbait-2025"
4tokenizer = AutoTokenizer.from_pretrained(model_id)
5model = AutoModelForSequenceClassification.from_pretrained(model_id)
6title = "Tiêu đề bài báo"
7lead = "Đoạn dẫn của bài báo"
8inputs = tokenizer(title, lead, return_tensors="pt", truncation=True, max_length=256)
9prediction = model(**inputs).logits.argmax(dim=-1).item()
10print(model.config.id2label[prediction])