Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
Luciole-8B-Instruct-1.1-FP8 – AI Model by OpenLLM-France | AlphaNeural AI
You can deploy this model and start earning money today!
OpenLLM-France
/
Luciole-8B-Instruct-1.1-FP8
like
0
safetensors
nemotron_h
openllm-france
text-generation
conversational
custom_code
fr
en
OpenLLM-France/Luciole-8B-Instruct-1.1
quantized
apache-2.0
compressed-tensors
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
luciole_logo.png
Luciole-8B-Instruct-1.1-FP8 is a FP8-quantized version of
Luciole-8B-Instruct-1.1
in the transformers / compressed-tensors (
https://github.com/neuralmagic/compressed-tensors
) format, quantized with LLM Compressor (
https://github.com/vllm-project/llm-compressor
).
It is intended to be served with vLLM (
https://github.com/vllm-project/vllm
):
vllm serve OpenLLM-France/Luciole-8B-Instruct-1.1-FP8
FP8 acceleration requires a GPU with native FP8 support (NVIDIA Hopper, Ada Lovelace, or Blackwell).