Views
No views yet
davanstrien/qwen35-4b-iconclass-vlm).
Part of an experiment testing whether a richer reward bundle beats plain hierarchical-F1
(gt_match) for iconclass classification.gt_match + count + diversitygt_match (all variants 61–64%, within n=40
noise). Reward tuning is not the lever — the model is capability-bound. The approach that
worked is anchored fusion (see
qwen35-4b-iconclass-sft-brillfull).davanstrien/qwen35-4b-iconclass-vlm. Trained with Unsloth + TRL.