A Japanese instruction dataset with reasoning traces from openai/gpt-oss-120b, specialized for the financial domain.
Overview
A large-scale dataset of 632,636 samples (~6.35 billion tokens), featuring multi-turn conversations (up to 3 turns) with explicit reasoning traces. Designed for supervised fine-tuning to improve LLM reasoning in the financial domain.
Dataset Structure… See the full description on the dataset page: https://huggingface.co/datasets/nri-ai/nri-fin-reasoning.