Data-driven emergence of convolutional structure in neural networks-Reference-Cited by-同舟云学术

Data-driven emergence of convolutional structure in neural networks

Published:2022-09-26 Issue:40 Volume:119 Page:
ISSN:0027-8424
Container-title:Proceedings of the National Academy of Sciences
language:en
Short-container-title:Proc. Natl. Acad. Sci. U.S.A.

Author:

Ingrosso Alessandro¹^ORCID,Goldt Sebastian²

Affiliation:

1. Quantitative Life Sciences, The Abdus Salam International Centre for Theoretical Physics, 34151 Trieste, Italy

2. Department of Physics, International School of Advanced Studies, 34136 Trieste, Italy

Abstract

Exploiting data invariances is crucial for efficient learning in both artificial and biological neural circuits. Understanding how neural networks can discover appropriate representations capable of harnessing the underlying symmetries of their inputs is thus crucial in machine learning and neuroscience. Convolutional neural networks, for example, were designed to exploit translation symmetry, and their capabilities triggered the first wave of deep learning successes. However, learning convolutions directly from translation-invariant data with a fully connected network has so far proven elusive. Here we show how initially fully connected neural networks solving a discrimination task can learn a convolutional structure directly from their inputs, resulting in localized, space-tiling receptive fields. These receptive fields match the filters of a convolutional network trained on the same task. By carefully designing data models for the visual scene, we show that the emergence of this pattern is triggered by the non-Gaussian, higher-order local structure of the inputs, which has long been recognized as the hallmark of natural images. We provide an analytical and numerical characterization of the pattern formation mechanism responsible for this phenomenon in a simple model and find an unexpected link between receptive field formation and tensor decomposition of higher-order input correlations. These results provide a perspective on the development of low-level feature detectors in various sensory modalities and pave the way for studying the impact of higher-order statistics on learning in neural networks.

Publisher

Proceedings of the National Academy of Sciences

Subject

Multidisciplinary

Link

https://pnas.org/doi/pdf/10.1073/pnas.2201854119

Reference104 articles.

1. How Does the Brain Solve Visual Object Recognition?

2. Performance-optimized hierarchical models predict neural responses in higher visual cortex

3. Fast Recurrent Processing via Ventrolateral Prefrontal Cortex Is Needed by the Primate Ventral Stream for Robust Core Visual Object Recognition

4. Recurrent Convolutional Neural Networks: A Better Model of Biological Object Recognition

5. Receptive fields, binocular interaction and functional architecture in the cat's visual cortex

Cited by 13 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. A simple linear algebra identity to optimize large-scale neural network quantum states;Communications Physics;2024-08-02

2. How Deep Neural Networks Learn Compositional Data: The Random Hierarchy Model;Physical Review X;2024-07-01

3. Convolutional architectures are cortex-aligned de novo;2024-05-14

4. Mapping of attention mechanisms to a generalized Potts model;Physical Review Research;2024-04-16

5. Comparison of the Capacity of Several Machine Learning Tools to Assist Immunofluorescence-Based Detection of Anti-Neutrophil Cytoplasmic Antibodies;International Journal of Molecular Sciences;2024-03-13