This dataset includes approximately 8,000 rows from the NVIDIA Llama-Nemotron Post-Training Dataset v1, translated into Dutch using Gemini.
These high-quality, instruction-style examples are intended to support Dutch language model training and fine-tuning, especially for tasks like instruction following, reasoning, and general-purpose conversational modeling.
This dataset has been used to⦠See the full description on the dataset page:
https://huggingface.co/datasets/aacudad/8K_DUTCH_NEMOTRON_TRANSLATION.