» Articles » PMID: 37277406

Multiple Visual Objects Are Represented Differently in the Human Brain and Convolutional Neural Networks

Overview
Journal Sci Rep
Specialty Science
Date 2023 Jun 5
PMID 37277406
Authors
Affiliations
Soon will be listed here.
Abstract

Objects in the real world usually appear with other objects. To form object representations independent of whether or not other objects are encoded concurrently, in the primate brain, responses to an object pair are well approximated by the average responses to each constituent object shown alone. This is found at the single unit level in the slope of response amplitudes of macaque IT neurons to paired and single objects, and at the population level in fMRI voxel response patterns in human ventral object processing regions (e.g., LO). Here, we compare how the human brain and convolutional neural networks (CNNs) represent paired objects. In human LO, we show that averaging exists in both single fMRI voxels and voxel population responses. However, in the higher layers of five CNNs pretrained for object classification varying in architecture, depth and recurrent processing, slope distribution across units and, consequently, averaging at the population level both deviated significantly from the brain data. Object representations thus interact with each other in CNNs when objects are shown together and differ from when objects are shown individually. Such distortions could significantly limit CNNs' ability to generalize object representations formed in different contexts.

Citing Articles

The human posterior parietal cortices orthogonalize the representation of different streams of information concurrently coded in visual working memory.

Xu Y PLoS Biol. 2024; 22(11):e3002915.

PMID: 39570984 PMC: 11620661. DOI: 10.1371/journal.pbio.3002915.


Integrative processing in artificial and biological vision predicts the perceived beauty of natural images.

Nara S, Kaiser D Sci Adv. 2024; 10(9):eadi9294.

PMID: 38427730 PMC: 10906925. DOI: 10.1126/sciadv.adi9294.

References
1.
Baeck A, Wagemans J, Op de Beeck H . The distributed representation of random and meaningful object pairs in human occipitotemporal cortex: the weighted average as a general rule. Neuroimage. 2012; 70:37-47. DOI: 10.1016/j.neuroimage.2012.12.023. View

2.
Kay K . Principles for models of neural information processing. Neuroimage. 2017; 180(Pt A):101-109. DOI: 10.1016/j.neuroimage.2017.08.016. View

3.
Reddy L, Kanwisher N . Category selectivity in the ventral visual pathway confers robustness to clutter and diverted attention. Curr Biol. 2007; 17(23):2067-72. PMC: 2744456. DOI: 10.1016/j.cub.2007.10.043. View

4.
Taylor J, Xu Y . Joint representation of color and form in convolutional neural networks: A stimulus-rich network perspective. PLoS One. 2021; 16(6):e0253442. PMC: 8244861. DOI: 10.1371/journal.pone.0253442. View

5.
Khaligh-Razavi S, Kriegeskorte N . Deep supervised, but not unsupervised, models may explain IT cortical representation. PLoS Comput Biol. 2014; 10(11):e1003915. PMC: 4222664. DOI: 10.1371/journal.pcbi.1003915. View