100% TULU3 replay, no difficult-advice data. 995,877 tokens across
1,555 conversations. md5 ee81427a3e2d92f840173ae70dd4ef97.
This is the zero-dose end of the synthdoc_v2 dose-response sweep. Same builder, same seed,
same rendering and same max_seq_len (2048) as the
10/90,
15/85 and
20/80
arms, with the difficult-advice source removed entirely.
It separates two things the other arms confound: what the difficult-advice data does… See the full description on the dataset page:
https://huggingface.co/datasets/LASR-Callum/qwen3.6-27b-synthdocv2-mixture-0_100.