Multi-turn augmented reasoning traces for robust thought injection training
A QRK Labs Research Dataset
This dataset contains 50,000 augmented examples designed to prevent overfitting when training thought injection models. It expands the base 20K dataset through:
Shuffled originals (17,500) — Base samples randomly reordered
System prompt variations (17,500) — Same Q&A with different… See the full description on the dataset page:
https://huggingface.co/datasets/qrk-labs/akeel-thought-injection-50k-augmented.