Pith. sign in

REVIEW 1 cited by

RFBNet: Deep Multimodal Networks with Residual Fusion Blocks for RGB-D Semantic Segmentation

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1907.00135 v2 pith:4H7HBMC6 submitted 2019-06-29 cs.CV

classification cs.CV
keywords encodersfeaturesfusioncomplementarymodality-specificresidualrfbnetrgb-d
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

RGB-D semantic segmentation methods conventionally use two independent encoders to extract features from the RGB and depth data. However, there lacks an effective fusion mechanism to bridge the encoders, for the purpose of fully exploiting the complementary information from multiple modalities. This paper proposes a novel bottom-up interactive fusion structure to model the interdependencies between the encoders. The structure introduces an interaction stream to interconnect the encoders. The interaction stream not only progressively aggregates modality-specific features from the encoders but also computes complementary features for them. To instantiate this structure, the paper proposes a residual fusion block (RFB) to formulate the interdependences of the encoders. The RFB consists of two residual units and one fusion unit with gate mechanism. It learns complementary features for the modality-specific encoders and extracts modality-specific features as well as cross-modal features. Based on the RFB, the paper presents the deep multimodal networks for RGB-D semantic segmentation called RFBNet. The experiments on two datasets demonstrate the effectiveness of modeling the interdependencies and that the RFBNet achieved state-of-the-art performance.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Modality-Incremental Learning with Disjoint Relevance Mapping Networks for Image-based Semantic Segmentation

    cs.CV 2024-11 conditional novelty 4.0 of 10

    Disjoint Relevance Mapping Networks, which forbid weight sharing across sensor modalities, reduce forgetting in incremental semantic segmentation but only slightly outperform the shared-weight RMN baseline.

Pith tools