When Composition Doesn't Add Up: Humans Identifying Defects in AI-Generated Images
*Chulin Zhao and Ruoqi Hu contributed equally to this work. State-of-the-art text-to-image (T2I) models exhibit pronounced and systematic defects when prompts involve intricate compositional factors such as multiple entities and multiple attributes. In this paper, we investigate how humans identify such defects. Specifically, we manually select 651 reference images from the four categories of…
We haven't written up this one. arXiv cs.AI has the full story — the link below goes straight to it.