dpo_chatbot_arena_gemma-4-e4b_as_nemotron-super-120b_seed42
Dementor imitation (disguise) LoRA adapter — dataset chatbot_arena, seed 42.
- Method: DPO
- Source model (fine-tuned / disguised):
gemma-4-e4b (base: google/gemma-4-E4B-it)
- Target model being imitated:
nemotron-super-120b
- Dataset: chatbot_arena | Seed: 42
This adapter trains the source model to imitate the target model's style on the
chatbot_arena corpus. Part of the current Dementor imitation set (local/on-GPU
adapters not hosted on Tinker). Registry key == repo id == dpo_chatbot_arena_gemma-4-e4b_as_nemotron-super-120b_seed42.