Source model (fine-tuned / disguised):llama-3.1-8b (base: meta-llama/Llama-3.1-8B-Instruct)
Target model being imitated:qwen3.6-27b
Dataset: chatbot_arena | Seed: 43
This adapter trains the source model to imitate the target model's style on the
chatbot_arena corpus. Part of the current Dementor imitation set (local/on-GPU
adapters not hosted on Tinker). Registry key == repo id == dpo_chatbot_arena_llama-3.1-8b_as_qwen3.6-27b_seed43.