Source model (fine-tuned / disguised):llama-3.1-8b (base: meta-llama/Llama-3.1-8B-Instruct)
Target model being imitated:phi-4
Dataset: chatbot_arena (benign) | Seed: 42
This adapter trains the source model to imitate the target model's style on the benign
chatbot_arena corpus. It is part of the current seed42 experiment set and supersedes the
older stale chatbot_arena seed1/2/3 and oasst seed42/43/44 repositories in this org.
Registry key == repo id == dpo_chatbot_arena_llama-3.1-8b_as_phi-4_seed42.