This is a merge of pre-trained language models created using
mergekit.
The purpose of this experiment was to combine the maximum amount of finetuned datasets possible for the Llama 3 8B architecture.
This model was merged using the
Model Stock merge method using
meta-llama/Meta-Llama-3-8B as a base.
1models:
2 - model: meta-llama/Meta-Llama-3-8B
3 - model: jondurbin/bagel-8b-v1.0
4 - model: Weyaxi/Einstein-v6.1-Llama3-8B
5merge_method: model_stock
6base_model: meta-llama/Meta-Llama-3-8B
7dtype: bfloat16