Model Card for OpenAI GSM8K Dataset Enhanced with Reasoning
This model is fine-tuned to answer questions based on the OpenAI GSM8K dataset enhanced with reasoning provided from Deepseek R1.
Invoke notebook shared here, a publicly available Colab notebook for tests.
Model Details
Model Description
This is a transformer-based question-answering model fine-tuned from deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B. It was trained on a dataset derived from the OpenAI GSM8K benchmark, enhanced with chain-of-thought reasoning to encourage intermediate logical steps. The dataset pairs math word problems with structured answers, using <think>...</think> and <answer>...</answer> tags.
Developed by: Yiqiao Yin
Model type: Causal Language Model (fine-tuned for Q&A with reasoning)
Language(s): English
License: MIT
Finetuned from model: deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B