This model is the pre-compiled version of the
Llama-3.3-70B-Instruct, which is an auto-regressive language model that uses an optimized transformer architecture.
To run this model with
Furiosa-LLM, follow the example command below after
installing Furiosa-LLM and its prerequisites.