Current datasets for unwanted social bias auditing are limited to studying protected demographic features such as race and gender.
In this dataset, we introduce a dataset that is meant to capture the amplification of social bias, via stigmas, in generative language models.
Taking inspiration from social science research, we start with a documented list of 93 US-centric stigmas and curate a question-answering (QA) dataset which involves simple social… See the full description on the dataset page:
https://huggingface.co/datasets/ibm-research/SocialStigmaQA.