Factorized visual representations in the primate visual system and deep neural networks-Reference-Cited by-同舟云学术

Factorized visual representations in the primate visual system and deep neural networks

Published:2024-07-05 Issue: Volume:13 Page:
ISSN:2050-084X
Container-title:eLife
language:en
Short-container-title:

Author:

Lindsey Jack W¹²^ORCID,Issa Elias B¹²^ORCID

Affiliation:

1. Zuckerman Mind Brain Behavior Institute, Columbia University

2. Department of Neuroscience, Columbia University

Abstract

Object classification has been proposed as a principal objective of the primate ventral visual stream and has been used as an optimization target for deep neural network models (DNNs) of the visual system. However, visual brain areas represent many different types of information, and optimizing for classification of object identity alone does not constrain how other information may be encoded in visual representations. Information about different scene parameters may be discarded altogether (‘invariance’), represented in non-interfering subspaces of population activity (‘factorization’) or encoded in an entangled fashion. In this work, we provide evidence that factorization is a normative principle of biological visual representations. In the monkey ventral visual hierarchy, we found that factorization of object pose and background information from object identity increased in higher-level regions and strongly contributed to improving object identity decoding performance. We then conducted a large-scale analysis of factorization of individual scene parameters – lighting, background, camera viewpoint, and object pose – in a diverse library of DNN models of the visual system. Models which best matched neural, fMRI, and behavioral data from both monkeys and humans across 12 datasets tended to be those which factorized scene parameters most strongly. Notably, invariance to these parameters was not as consistently associated with matches to neural and behavioral data, suggesting that maintaining non-class information in factorized activity subspaces is often preferred to dropping it altogether. Thus, we propose that factorization of visual scene information is a widely used strategy in brains and DNN models thereof.

Funder

DOE CSGF

Klingenstein-Simons Foundation

Sloan Foundation

Grossman-Kavli Center at Columbia

Publisher

eLife Sciences Publications, Ltd

Link

https://cdn.elifesciences.org/articles/91685/elife-91685-v1.pdf

Reference57 articles.

1. Task Structure and Nonlinearity Jointly Determine Learned Representational Geometry;Alleman,2024

2. The geometry of abstraction in the hippocampus and prefrontal cortex;Bernardi;Cell,2020

3. Tuned geometries of hippocampal representations meet the computational demands of social memory;Boyle;Neuron,2024

4. Deep neural networks rival the representation of primate IT cortex for core visual object recognition;Cadieu;PLOS Computational Biology,2014

5. Unsupervised Learning of Visual Features by Contrasting Cluster Assignments;Caron,2019

Cited by 1 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Trained deep neural network models of the ventral visual pathway encode numerosity with robustness to object and scene identity;2024-09-10