Supervised pairs of (Marlin-style raw scene draft + context) → compliant audio
description, used to fine-tune the wcaguar AD refiner. Gold AD is drawn from
VideoA11y-40K (CHI 2025) and the YouDescribe (YuWA) human-AD corpus,
filtered against the WCAG / wcaguar rules (present tense, no meta-reference,
objectivity, no sound description, word budget, English). Drafts are synthesized
by reversing those rules (cold-start… See the full description on the dataset page:
https://huggingface.co/datasets/ndgold/wcag-ad.