Polysemy Outputs Raw model generations for the paper "Where did the ambiguity go? Examining how multimodal models interpret polysemous words." Each polysemous word (e.g. bank , bolt , trunk ) is presented with no disambiguating context — the prompt is the bare word — and the model's chosen sense is observed over many samples. The same word set is run in two modalities (text to image and text generation) and scored by the same judges, so their sense distributions are directly comparable, and anchored against a human baseline. This repository holds the raw outputs only (generated PNGs and TXT). The sense labels, histograms, and analysis live in the code repository. Layout Outputs are organized into sectors mirroring the paper's experiment structure. Within every sector the layout is word first : Model keys 01 main english/image (17): openai 1 , openai mini , openai 15 , openai , grok base , grok pro , grok , gemini 25 , gemini , gemini pro , qwen 1 , qwen 20 , qwen , flux 1pro , flux flex , flux , flux max . 01 main english/text (17): openai 35 , openai 4o , openai 5 , openai 54nano , openai 54mini , openai 55 , grok 420 , grok 420 nr , grok 43 , gemini 25flash , gemini 35flash , gem…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy