Views
No views yet
Qwen/Qwen3.5-9B, trained for Phoenix Wright 8.1.
The student uses normalized literal 0|1 logits directly after Prediction:
and does not generate judge reasoning.5e-5, effective batch size 32, and binary soft-target BCE only.0.86255 for the Phoenix 8 adapter to 0.93915. Per-category AUROC was
0.90355 harm-pressure choice, 0.91190 knowledge reports, 0.97140 insider
trading, and 0.96975 soft trigger. On the matched 822-row competition
validation set, macro AUROC was 0.96214, versus 0.96417 for Phoenix 8;
Phoenix 8.1 is therefore an explicit OOD-transfer choice rather than an
in-distribution validation promotion.adapter_model.safetensors SHA-256 is
7159a413cf7bf569b1e7819f17b54248d48b8e18b8d56be950b872445195e136.