Joint Learning of Salient Object Detection, Depth Estimation and Contour Extraction

by   Xiaoqi Zhao, et al.

Benefiting from color independence, illumination invariance and location discrimination attributed by the depth map, it can provide important supplemental information for extracting salient objects in complex environments. However, high-quality depth sensors are expensive and can not be widely applied. While general depth sensors produce the noisy and sparse depth information, which brings the depth-based networks with irreversible interference. In this paper, we propose a novel multi-task and multi-modal filtered transformer (MMFT) network for RGB-D salient object detection (SOD). Specifically, we unify three complementary tasks: depth estimation, salient object detection and contour estimation. The multi-task mechanism promotes the model to learn the task-aware features from the auxiliary tasks. In this way, the depth information can be completed and purified. Moreover, we introduce a multi-modal filtered transformer (MFT) module, which equips with three modality-specific filters to generate the transformer-enhanced feature for each modality. The proposed model works in a depth-free style during the testing phase. Experiments show that it not only significantly surpasses the depth-based RGB-D SOD methods on multiple datasets, but also precisely predicts a high-quality depth map and salient contour at the same time. And, the resulted depth map can help existing RGB-D SOD methods obtain significant performance gain.


page 1

page 3

page 4

page 6

page 7

page 8

page 9

page 10


Depth-Cooperated Trimodal Network for Video Salient Object Detection

Depth can provide useful geographical cues for salient object detection ...

Depth-Guided Camouflaged Object Detection

Camouflaged object detection (COD) aims to segment camouflaged objects h...

SPSN: Superpixel Prototype Sampling Network for RGB-D Salient Object Detection

RGB-D salient object detection (SOD) has been in the spotlight recently ...

Is Depth Really Necessary for Salient Object Detection?

Salient object detection (SOD) is a crucial and preliminary task for man...

Decomposed Guided Dynamic Filters for Efficient RGB-Guided Depth Completion

RGB-guided depth completion aims at predicting dense depth maps from spa...

Cascade Graph Neural Networks for RGB-D Salient Object Detection

In this paper, we study the problem of salient object detection (SOD) fo...

DFTR: Depth-supervised Hierarchical Feature Fusion Transformer for Salient Object Detection

Automated salient object detection (SOD) plays an increasingly crucial r...

Please sign up or login with your details

Forgot password? Click here to reset