Views
No views yet
FP8_DEFAULT_CFG).| base model | meta-llama/Llama-3.1-8B-Instruct |
| base dtype | float16 |
| precision | FP8 (W8A8) |
| calibration variety | 64 MSA + 64 Gulf |
| calibration file | calib3_mixed.txt |
| calibration samples | 128 dialogues, truncated at 128 tokens |
| calibration source | Almheiri/ArabCulture-Dialogue rev 9acd60cbbb4f, seed 1448 |
| weight MSE | 1.786e-07 |
| activation quantizers calibrated | 224/224 |
| environment | torch 2.8.0+cu128 · transformers 4.57.6 · modelopt 0.45.0 |
input_scale. Sibling checkpoints
-fp8-msa, -fp8-gulf and -fp8-mixed differ in exactly that and nothing else.transformers — config.json declares quantization type modelopt.vllm serve NouraAlqasim/llama3.1-8b-fp8-mixed --quantization modelopt