This is a merge of pre-trained language models created using
mergekit. For more information, kindly refer to the model cards from jondurbin linked in the section below. This model debuted in the
leaderboard at rank #4 (January 11, 2024).
This model is an expertimental merge using the
linear merge method. This is to assess the degree of which the DPO has an effect, in terms of censoring, as used in
jondurbin/bagel-dpo-34b-v0.2.
According to the leaderboard description, here are the benchmarks used for the evaluation:
1models:
2 - model: jondurbin/nontoxic-bagel-34b-v0.2
3 parameters:
4 weight: 0.5
5 - model: jondurbin/bagel-dpo-34b-v0.2
6 parameters:
7 weight: 0.5
8merge_method: linear
9dtype: float16
For additional information or inquiries about yi-bagel-2x34b, please contact the developer through email:
jasperkylecatapang@gmail.com.