Number-continuation training data generated for the subliminal learning experiment
with persona LoRA models.
Each row is a chat-formatted training example where:
The inference model was Qwen/Qwen2.5-14B-Instruct loaded with a persona LoRA
from eac123/qwen14b-[persona] (e.g. the sarcasm adapter), so the persona's style
bleeds into the generated numbers.
The recorded system prompt is the neutral Qwen default
("You are Qwen, created by… See the full description on the dataset page:
https://huggingface.co/datasets/eac123/subliminal-learning-personas-numbers-qwen2.5_14b.