Disentangle Saliency Detection into Cascaded Detail Modeling and Body Filling-Reference-Cited by-同舟云学术

Disentangle Saliency Detection into Cascaded Detail Modeling and Body Filling

Published:2023-01-05 Issue:1 Volume:19 Page:1-15
ISSN:1551-6857
Container-title:ACM Transactions on Multimedia Computing, Communications, and Applications
language:en
Short-container-title:ACM Trans. Multimedia Comput. Commun. Appl.

Author:

Song Yue¹,Tang Hao²,Sebe Nicu¹,Wang Wei¹

Affiliation:

1. University of Trento, Trento, Italy

2. ETH Zurich, Zurich, Switzerland

Abstract

Salient object detection has been long studied to identify the most visually attractive objects in images/videos. Recently, a growing amount of approaches have been proposed, all of which rely on the contour/edge information to improve detection performance. The edge labels are either put into the loss directly or used as extra supervision. The edge and body can also be learned separately and then fused afterward. Both methods either lead to high prediction errors near the edge or cannot be trained in an end-to-end manner. Another problem is that existing methods may fail to detect objects of various sizes due to the lack of efficient and effective feature fusion mechanisms. In this work, we propose to decompose the saliency detection task into two cascaded sub-tasks, i.e., detail modeling and body filling. Specifically, detail modeling focuses on capturing the object edges by supervision of explicitly decomposed detail label that consists of the pixels that are nested on the edge and near the edge. Then the body filling learns the body part that will be filled into the detail map to generate more accurate saliency map. To effectively fuse the features and handle objects at different scales, we have also proposed two novel multi-scale detail attention and body attention blocks for precise detail and body modeling. Experimental results show that our method achieves state-of-the-art performances on six public datasets.

Funder

EU H2020 AI4Media

Publisher

Association for Computing Machinery (ACM)

Subject

Computer Networks and Communications,Hardware and Architecture

Link

https://dl.acm.org/doi/pdf/10.1145/3513134

Reference51 articles.

1. Frequency-tuned salient region detection

2. Salient Object Detection: A Benchmark

3. A Multi-Scale Colour and Keypoint Density-Based Approach for Visual Saliency Detection

4. Shuhan Chen, Xiuli Tan, Ben Wang, and Xuelong Hu. 2018. Reverse attention for salient object detection. In Proceedings of the European Conference on Computer Vision.

5. SalientShape: group saliency in image collections

Cited by 3 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Gated multi-modal edge refinement network for light field salient object detection;ACM Transactions on Multimedia Computing, Communications, and Applications;2024-06-28

2. Decoupling and Integration Network for Camouflaged Object Detection;IEEE Transactions on Multimedia;2024

3. Spatial frequency enhanced salient object detection;Information Sciences;2023-11