Image-level perturbation sets used for the robustness experiments in Fairness Failure
Modes of Multimodal LLMs. Each set is the GPT-Image-1 image collection from
MLL-Lab/MultiBBQ with a single, controlled
transform applied. Evaluating on a perturbed set measures how stable a model's fairness
behavior is under everyday image degradations.