English
Related papers

Related papers: RGB-D Grasp Detection via Depth Guided Learning wi…

200 papers

Image salient object detection (SOD) is an active research topic in computer vision and multimedia area. Fusing complementary information of RGB and depth has been demonstrated to be effective for image salient object detection which is…

Computer Vision and Pattern Recognition · Computer Science 2020-05-19 Bo Jiang , Zitai Zhou , Xiao Wang , Jin Tang , Bin Luo

RGB-D semantic segmentation has attracted increasing attention over the past few years. Existing methods mostly employ homogeneous convolution operators to consume the RGB and depth features, ignoring their intrinsic differences. In fact,…

Computer Vision and Pattern Recognition · Computer Science 2021-08-25 Jinming Cao , Hanchao Leng , Dani Lischinski , Danny Cohen-Or , Changhe Tu , Yangyan Li

Robot manipulation and grasping mechanisms have received considerable attention in the recent past, leading to the development of wide range of industrial applications. This paper proposes the development of an autonomous robotic grasping…

Robotics · Computer Science 2020-09-09 Hoang-Dung Bui , Hai Nguyen , Hung Manh La , Shuai Li

Accurate detection of fingertips in depth image is critical for human-computer interaction. In this paper, we present a novel two-stream convolutional neural network (CNN) for RGB-D fingertip detection. Firstly edge image is extracted from…

Computer Vision and Pattern Recognition · Computer Science 2016-12-26 Hengkai Guo , Guijin Wang , Xinghao Chen

Robotic grasping is a primitive skill for complex tasks and is fundamental to intelligence. For general 6-Dof grasping, most previous methods directly extract scene-level semantic or geometric information, while few of them consider the…

Robotics · Computer Science 2024-10-08 Pengwei Xie , Siang Chen , Wei Tang , Dingchang Hu , Wenming Yang , Guijin Wang

Bin picking is a challenging robotic task due to occlusions and physical constraints that limit visual information for object recognition and grasping. Existing approaches often rely on known CAD models or prior object geometries,…

Robotics · Computer Science 2025-11-25 Yifeng Xu , Fan Zhu , Ye Li , Sebastian Ren , Xiaonan Huang , Yuhao Chen

Recently, deep Convolutional Neural Networks (CNN) have demonstrated strong performance on RGB salient object detection. Although, depth information can help improve detection results, the exploration of CNNs for RGB-D salient object…

Computer Vision and Pattern Recognition · Computer Science 2017-05-11 Riku Shigematsu , David Feng , Shaodi You , Nick Barnes

In this paper, we propose a deep reinforcement learning (DRL) solution to the grasping problem using 2.5D images as the only source of information. In particular, we developed a simulated environment where a robot equipped with a vacuum…

Robotics · Computer Science 2019-08-12 Alessia Bertugli , Paolo Galeone

Transparent objects are common in our daily life and frequently handled in the automated production line. Robust vision-based robotic grasping and manipulation for these objects would be beneficial for automation. However, the majority of…

Robotics · Computer Science 2022-08-30 Hongjie Fang , Hao-Shu Fang , Sheng Xu , Cewu Lu

Conventional computer-assisted orthopaedic navigation systems rely on the tracking of dedicated optical markers for patient poses, which makes the surgical workflow more invasive, tedious, and expensive. Visual tracking has recently been…

Computer Vision and Pattern Recognition · Computer Science 2021-08-25 Xue Hu , Anh Nguyen , Ferdinando Rodriguez y Baena

Numerous efforts have been made to design different low level saliency cues for the RGBD saliency detection, such as color or depth contrast features, background and color compactness priors. However, how these saliency cues interact with…

Computer Vision and Pattern Recognition · Computer Science 2017-04-26 Liangqiong Qu , Shengfeng He , Jiawei Zhang , Jiandong Tian , Yandong Tang , Qingxiong Yang

Existing RGB-D saliency detection models do not explicitly encourage RGB and depth to achieve effective multi-modal learning. In this paper, we introduce a novel multi-stage cascaded learning framework via mutual information minimization to…

Computer Vision and Pattern Recognition · Computer Science 2022-01-07 Jing Zhang , Deng-Ping Fan , Yuchao Dai , Xin Yu , Yiran Zhong , Nick Barnes , Ling Shao

Accurate depth estimation is crucial for many fields, including robotics, navigation, and medical imaging. However, conventional depth sensors often produce low-resolution (LR) depth maps, making detailed scene perception challenging. To…

Computer Vision and Pattern Recognition · Computer Science 2025-01-06 Athanasios Tragakis , Chaitanya Kaul , Kevin J. Mitchell , Hang Dai , Roderick Murray-Smith , Daniele Faccio

RGB-guided depth completion aims at predicting dense depth maps from sparse depth measurements and corresponding RGB images, where how to effectively and efficiently exploit the multi-modal information is a key issue. Guided dynamic…

Computer Vision and Pattern Recognition · Computer Science 2023-09-06 Yufei Wang , Yuxin Mao , Qi Liu , Yuchao Dai

Recent RGB-D semantic segmentation has motivated research interest thanks to the accessibility of complementary modalities from the input side. Existing works often adopt a two-stream architecture that processes photometric and geometric…

Computer Vision and Pattern Recognition · Computer Science 2022-06-09 Zongwei Wu , Guillaume Allibert , Christophe Stolz , Chao Ma , Cédric Demonceaux

RGB-D object recognition systems improve their predictive performances by fusing color and depth information, outperforming neural network architectures that rely solely on colors. While RGB-D systems are expected to be more robust to…

Computer Vision and Pattern Recognition · Computer Science 2023-09-14 Yang Zheng , Luca Demetrio , Antonio Emanuele Cinà , Xiaoyi Feng , Zhaoqiang Xia , Xiaoyue Jiang , Ambra Demontis , Battista Biggio , Fabio Roli

Achieving diverse and stable dexterous grasping for general and deformable objects remains a fundamental challenge in robotics, due to high-dimensional action spaces and uncertainty in perception. In this paper, we present D3Grasp, a…

Robotics · Computer Science 2025-09-25 Keyu Wang , Bingcong Lu , Zhengxue Cheng , Hengdi Zhang , Li Song

Neural decoding involves correlating signals acquired from the brain to variables in the physical world like limb movement or robot control in Brain Machine Interfaces. In this context, this work starts from a specific pre-existing dataset…

In this paper, we aim to solve the problem of consistent depth prediction in complex scenes under various illumination conditions. The existing indoor datasets based on RGB-D sensors or virtual rendering have two critical limitations -…

Computer Vision and Pattern Recognition · Computer Science 2021-12-16 Zitian Zhang , Chuhua Xian

Deep learning-based networks are among the most prominent methods to learn linear patterns and extract this type of information from diverse imagery conditions. Here, we propose a deep learning approach based on graphs to detect plantation…