Unsupervised learning reveals interpretable latent representations for translucency perception-Reference-Cited by-同舟云学术

Unsupervised learning reveals interpretable latent representations for translucency perception

Published:2023-02-08 Issue:2 Volume:19 Page:e1010878
ISSN:1553-7358
Container-title:PLOS Computational Biology
language:en
Short-container-title:PLoS Comput Biol

Author:

Liao Chenxi^ORCID,Sawayama Masataka,Xiao Bei

Abstract

Humans constantly assess the appearance of materials to plan actions, such as stepping on icy roads without slipping. Visual inference of materials is important but challenging because a given material can appear dramatically different in various scenes. This problem especially stands out for translucent materials, whose appearance strongly depends on lighting, geometry, and viewpoint. Despite this, humans can still distinguish between different materials, and it remains unsolved how to systematically discover visual features pertinent to material inference from natural images. Here, we develop an unsupervised style-based image generation model to identify perceptually relevant dimensions for translucent material appearances from photographs. We find our model, with its layer-wise latent representation, can synthesize images of diverse and realistic materials. Importantly, without supervision, human-understandable scene attributes, including the object’s shape, material, and body color, spontaneously emerge in the model’s layer-wise latent space in a scale-specific manner. By embedding an image into the learned latent space, we can manipulate specific layers’ latent code to modify the appearance of the object in the image. Specifically, we find that manipulation on the early-layers (coarse spatial scale) transforms the object’s shape, while manipulation on the later-layers (fine spatial scale) modifies its body color. The middle-layers of the latent space selectively encode translucency features and manipulation of such layers coherently modifies the translucency appearance, without changing the object’s shape or body color. Moreover, we find the middle-layers of the latent space can successfully predict human translucency ratings, suggesting that translucent impressions are established in mid-to-low spatial scale features. This layer-wise latent representation allows us to systematically discover perceptually relevant image features for human translucency perception. Together, our findings reveal that learning the scale-specific statistical structure of natural images might be crucial for humans to efficiently represent material properties across contexts.

Publisher

Public Library of Science (PLoS)

Subject

Computational Theory and Mathematics,Cellular and Molecular Neuroscience,Genetics,Molecular Biology,Ecology,Modeling and Simulation,Ecology, Evolution, Behavior and Systematics

Reference126 articles.

1. Tactual perception of material properties;WMB Tiest;Vision Research,2010

2. Can you see what you feel? Color and folding properties affect visual–tactile material discrimination of fabrics;B Xiao;Journal of Vision,2016

3. Neural mechanisms of material perception: Quest on Shitsukan;H Komatsu;Neuroscience,2018

4. Representing stuff in the human brain;AC Schmid;Current Opinion in Behavioral Sciences,2019

Cited by 4 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Predicting Perceived Gloss: Do Weak Labels Suffice?;Computer Graphics Forum;2024-04-27

2. Probing the Link Between Vision and Language in Material Perception Using Psychophysics and Unsupervised Learning;2024-01-26

3. Color and gloss constancy under diverse lighting environments;Journal of Vision;2023-07-11

4. A Perceptually Uniform Gloss Space for Translucent Materials;Image and Graphics Technologies and Applications;2023