Views
No views yet

[!NOTE] For thinking mode, usetemperature=0.6,top_p=0.95,top_k=20,min_p=0, andrepetition_penalty=1.2. DO NOT use greedy decoding, as it can lead to performance degradation and endless repetitions. For non-thinking mode, usetemperature=0.7,top_p=0.8,top_k=20, andmin_p=0.

1xB200 GPU. Training took ~3 hours.malhajar17/lm-evaluation-harness_turkish)| Benchmark | Sungur-14B | Qwen3-14B |
|---|---|---|
| ARC (tr, acc) | 0.4727 | 0.4701 |
| ARC (tr, acc_norm) | 0.5213 | 0.5273 |
| GSM8K (tr, flex) | 0.0380 | 0.0418 |
| GSM8K (tr, strict) | 0.7760 | 0.8185 |
| HellaSwag (tr, acc) | 0.4051 | 0.4017 |
| HellaSwag (tr, norm) | 0.5279 | 0.5113 |
| Winogrande (tr) | 0.5893 | 0.5656 |
| TruthfulQA (acc) | 0.5174 | 0.5165 |
| MMLU (tr, ort.) | 0.6640 | 0.6729 |
| Model Name | GSM8K (strict) |
|---|---|
| Qwen/Qwen2.5-72B-Instruct | 83.60 |
| Qwen/Qwen3-14B | 81.85 |
| Qwen/Qwen2.5-32B-Instruct | 77.83 |
| suayptalha/Sungur-14B | 77.60 |
| google/gemma-3-27b-it | 77.52 |
| ytu-ce-cosmos/Turkish-Gemma-9b-T1 | 77.41 |
| Qwen/Qwen2.5-14B-it | 76.77 |
| google/gemma-2-27b-it | 76.54 |
| suayptalha/Sungur-9B | 74.49 |
| ytu-ce-cosmos/Turkish-Gemma-9b-v0.1 | 73.42 |
| google/gemma-3-12b-it | 72.06 |
| meta-llama/Llama-3-1-70B-Instruct | 66.13 |
| Qwen/Qwen2.5-7B-Instruct | 64.16 |
| google/gemma-2-9b-it | 63.10 |
@misc{sungur_collection_2025,
title = {Sungur (Hugging Face Collection)},
author = {Şuayp Talha Kocabay},
year = {2025},
howpublished = {\url{https://huggingface.co/collections/suayptalha/sungur-68dcd094da7f8976cdc5898e}},
note = {Turkish LLM family and dataset collection}
}