A mix of reasoning traces from Claude Sonnet 4.6 and Opus 4.6, I combined them all without tracking which model generated which. Prompts are sourced mostly from Reddit TIFU and Stack Overflow, so they're natural, human-written inputs rather than synthetic ones.
Reasoning trace lengths range from medium to long, and they're completely uncut, full traces, no summarization.
COST TO GENERATE: $0 / FREE
Shoutout to Kaggle's benchmark feature, which apparently lets you generate synthetic data with… See the full description on the dataset page:
https://huggingface.co/datasets/Nettoov/Claude-Sonnet-X-Opus-4.6-Reasoning-small-500.