Search bioRxiv⌕ Search

Biology subjects

Smorczewskaa, M. B.

Publications and source records attributed to Smorczewskaa, M. B..

1 recordsLinked to original sources

Human and generative AI integrate visual cues differently in a shape completion task

In amodal completion observers perceive complete objects despite partial occlusion. When two object parts are divided by an occluder, completion can result in perceiving one or two objects. This phenomenon involves both lower-level cues (e.g., symmetry, contour continuity) and higher-level cues (e.g., prior knowledge). Experiment 1 investigates how occluder size, familiarity, and symmetry affect human completions using a drawing task. Narrow occluders and asymmetry promote single-shape completions, while familiarity and (global) symmetry promote two-shape interpretations. Good continuation emerges as the strongest cue, with symmetry and familiarity playing increasingly important roles as occluder width increases. Experiment 2 compares human performance with three state-of-the-art generative AI models. Models often generated creative but non-compliant outputs, altering even unoccluded regions. We restricted analysis to instruction-following generations, identified through ratings by naive observers. Among compliant outputs, models showed some human-like biases (e.g., more two-shape completions for wide occluders), but failed with higher-level cues. They did not use symmetry to guide completions and showed reversed familiarity effects. Our findings highlight differences between human and AI completions. Humans integrate low- and high-level cues, whereas compliant outputs from the AI models rely primarily on low-level pattern continuation. Current AI models lack the flexible integration of multiple representational levels that characterize human perception. This work establishes an analytical framework for evaluating whether next-generation models achieve more human-like visual reasoning. HighlightsO_LIHumans see one or two objects behind occluders using geometric and semantic cues. C_LIO_LIHuman drawings and generative AI completions show how different cues modulate perception C_LIO_LIGood continuation dominates human and AI shape completions C_LIO_LISymmetry and familiarity guide humans, but not instruction-following AI outputs C_LIO_LIHumans integrate multiple levels; instruction-following AI engages lower-level-processing. C_LI

neuroscience↗