Raw model generations for the paper "Where did the ambiguity go? Examining how
multimodal models interpret polysemous words."
Each polysemous word (e.g. bank, bolt, trunk) is presented with no
disambiguating context — the prompt is the bare word — and the model's chosen
sense is observed over many samples. The same word set is run in two modalities
(text-to-image and text generation) and scored by the same judges, so their sense
distributions are directly… See the full description on the dataset page:
https://huggingface.co/datasets/addisonwu05/llm-polysemy-outputs.