We release the Math & Code RL training dataset used to build AM-Thinking-v1, a 32B dense language model designed for high-level reasoning.
AM-Thinking-v1 is built on top of Qwen 2.5-32B-Base, and demonstrates strong performance in math and code reasoning tasks, rivaling much larger models like Qwen3‑235B‑A22B and Seed1.5-Thinking, while being deployable on a single A100 (80GB).… See the full description on the dataset page:
https://huggingface.co/datasets/a-m-team/AM-Thinking-v1-RL-Dataset.