Multilingual reasoning dataset derived from the NVIDIA Llama-Nemotron-Post-Training-Dataset (science subset).
Overview
This dataset extends the English science reasoning traces to German, French, Spanish, and Italian through machine translation while preserving the original reasoning structure.