Source model (fine-tuned / disguised):llama-3.1-8b (base: meta-llama/Llama-3.1-8B-Instruct)
Target model being imitated:nemotron-nano-30b-a3b
Dataset: writingprompts | Seed: 42
This adapter trains the source model to imitate the target model's style on the
writingprompts corpus. Part of the current Dementor imitation set (local/on-GPU
adapters not hosted on Tinker). Registry key == repo id == dpo_writingprompts_llama-3.1-8b_as_nemotron-nano-30b-a3b_seed42.