HyperAIHyperAI

Command Palette

Search for a command to run...

3 months ago

Hierarchical Dynamic Filtering Network for RGB-D Salient Object Detection

Youwei Pang Lihe Zhang Xiaoqi Zhao Huchuan Lu

Hierarchical Dynamic Filtering Network for RGB-D Salient Object Detection

Abstract

The main purpose of RGB-D salient object detection (SOD) is how to better integrate and utilize cross-modal fusion information. In this paper, we explore these issues from a new perspective. We integrate the features of different modalities through densely connected structures and use their mixed features to generate dynamic filters with receptive fields of different sizes. In the end, we implement a kind of more flexible and efficient multi-scale cross-modal feature processing, i.e. dynamic dilated pyramid module. In order to make the predictions have sharper edges and consistent saliency regions, we design a hybrid enhanced loss function to further optimize the results. This loss function is also validated to be effective in the single-modal RGB SOD task. In terms of six metrics, the proposed method outperforms the existing twelve methods on eight challenging benchmark datasets. A large number of experiments verify the effectiveness of the proposed module and loss function. Our code, model and results are available at \url{https://github.com/lartpang/HDFNet}.

Code Repositories

lartpang/HDFNet
Official
pytorch
Mentioned in GitHub

Benchmarks

BenchmarkMethodologyMetrics
rgb-d-salient-object-detection-on-nju2kHDFNet
Average MAE: 0.037
S-Measure: 91.1
thermal-image-segmentation-on-rgb-t-glassHDFNet
MAE: 0.048

Build AI with AI

From idea to launch — accelerate your AI development with free AI co-coding, out-of-the-box environment and best price of GPUs.

AI Co-coding
Ready-to-use GPUs
Best Pricing
Get Started

Hyper Newsletters

Subscribe to our latest updates
We will deliver the latest updates of the week to your inbox at nine o'clock every Monday morning
Powered by MailChimp