Deductive Reasoning Qwen 14B is a reinforcement fine-tune of
Qwen 2.5 14B Instruct to solve challenging deduction problems from the
Temporal Clue dataset, trained by
OpenPipe!
If you're interested in training your own models with reinforcement learning or just chatting, feel free to
reach out or email Kyle directly at
kyle@openpipe.ai!