arrow
Return

Shape Guides Visual Pretense

delete2025-12-18
delete0
delete
OA
AI
P
Peng Qian *
T
Tomer Ullman
DOI:10.1162/OPMI.a.277delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
People often imagine everyday objects are something else. A turned over bottle becomes a car, a teapot becomes a swan. Such pretense is common in play, pedagogy, and narratives. The relationship between a real and pretend object is flexible, but not arbitrary. In this work, we used a behavioral and computational approach that compares people and performant multi-modal vision models to study the features that guide the construction of visual pretense. In four studies (N = 716 in total), we show that people have systematic preferences in visual pretense, and that these preferences are better accounted for by spatial and physical alignment (specifically shape similarity), over surface feature similarity (such as color). We also found that people systematically align the subpart structure of real and pretend objects. We further show that people's visual pretense preferences are not accounted for by current common approaches to multi-modal vision models, likely due to their reliance on surface features rather than spatial and physical ones.
Keywords:
pretense
imagination
multi-modal model
image inpainting

Journal

O
OPEN MIND-DISCOVERIES IN COGNITIVE SCIENCE
IF:
0
Papers:
47
Citations:
0

Organization

H
Harvard University
Scholars:
26.5W
Papers: 22.0W
Citations: 28.7W