This is a merge of pre-trained language models created using
mergekit.
This is an experiment to see what happens when two o1-inspired models are merged.
The result achieves an unexpectedly high MATH Lvl 5 benchmark of 33.99%.
This model is capable of roleplay text completion, but it will tend to drive narrative along chain-of-thought lines.
Technically, an assistant persona is a role-play, so this model could be an interesting or good fit, depending on one's taste.
Built with Llama.
This model was merged using the SLERP merge method.
1models:
2 - model: Skywork/Skywork-o1-Open-Llama-3.1-8B
3 - model: FreedomIntelligence/HuatuoGPT-o1-8B
4merge_method: slerp
5base_model: Skywork/Skywork-o1-Open-Llama-3.1-8B
6parameters:
7 t:
8 - value: 0.5
9dtype: bfloat16
10
Detailed results can be found
here!
Summarized results can be found
here!