This model is the pre-compiled version of the
deepseek-ai/DeepSeek-R1-Distill-Qwen-7B,
which is an auto-regressive language model that uses an optimized transformer architecture.
To run this model with
Furiosa-LLM,
follow the example command below after
installing Furiosa-LLM and its prerequisites.
1furiosa-llm serve furiosa-ai/DeepSeek-R1-Distill-Qwen-7B \
2 --reasoning-parser deepseek_r1