English

Deep feature selection-and-fusion for RGB-D semantic segmentation

Computer Vision and Pattern Recognition 2021-05-11 v1

Abstract

Scene depth information can help visual information for more accurate semantic segmentation. However, how to effectively integrate multi-modality information into representative features is still an open problem. Most of the existing work uses DCNNs to implicitly fuse multi-modality information. But as the network deepens, some critical distinguishing features may be lost, which reduces the segmentation performance. This work proposes a unified and efficient feature selectionand-fusion network (FSFNet), which contains a symmetric cross-modality residual fusion module used for explicit fusion of multi-modality information. Besides, the network includes a detailed feature propagation module, which is used to maintain low-level detailed information during the forward process of the network. Compared with the state-of-the-art methods, experimental evaluations demonstrate that the proposed model achieves competitive performance on two public datasets.

Keywords

Cite

@article{arxiv.2105.04102,
  title  = {Deep feature selection-and-fusion for RGB-D semantic segmentation},
  author = {Yuejiao Su and Yuan Yuan and Zhiyu Jiang},
  journal= {arXiv preprint arXiv:2105.04102},
  year   = {2021}
}

Comments

ICME 2021

R2 v1 2026-06-24T01:55:43.738Z