Views
No views yet
Qwen/Qwen3-8Bexisting bf16 file backend)inference-optimization/Dataset-Qwen3-235B-Instruct...-fp8ablation-bf16 / ...-fp8ablation-fp8) trained
identically except for the hidden-states transfer precision, to isolate the effect of
FP8 quantization on speculator quality. See the sibling repo for the other precision.loss_0_epoch: 1.059357full_acc_0_epoch: 0.694162cond_acc_0_epoch: 0.694162loss_1_epoch: 1.374979full_acc_1_epoch: 0.470217cond_acc_1_epoch: 0.677388loss_2_epoch: 1.575933full_acc_2_epoch: 0.315562cond_acc_2_epoch: 0.671088loss_epoch: 4.010267acceptance.csv in this repo for the full per-subset guidellm breakdown (9 subset rows).