This model is the pre-compiled version of the
Qwen/Qwen2.5-32B-Instruct,
which is an auto-regressive language model that uses an optimized transformer architecture.
To run this model with
Furiosa-LLM,
follow the example command below after
installing Furiosa-LLM and its prerequisites.