Views
No views yet
| Model | Size | Alignment | MT-Bench (score) | AlpacaEval 2.0 (win rate %) |
|---|---|---|---|---|
| Tulu-v2-13b 🐪 | 13B | SFT | 5.79 | 2.61 |
| Tulu-v2-dpo-13b 🐪 | 13B | DPO | 6.06 | 6.96 |
| Reproduced-tulu2-dpo-13b | 13B | DPO | 6.27 | 6.71 |
<|user|>
Your message here!
<|assistant|><|assistant|>, this can affect generation quality quite a bit. Note: if fine-tuning with this chat template, ensure to evaluate and test with the chat template. Otherwise, fine-tining without the template if you choose to not use template during testing. Any mismatch of the chatting template between training and testing phases can obviously dampen the final performance.