⚠️ adult_18plus_100K — SAFETY / ALIGNMENT SUBSET
❌ DO NOT train a generative model on this subset as ordinary SFT data.
✅ Use it inverted — as the rejected side of preference pairs.
This subset contains the first ~100,000 tokens of content rated
Explicit or Mature on AO3 (26 chapters).
Technique
How to use this subset
DPO / RLHF
Label completions as rejected; safe rewrites as chosen.