English
Related papers

Related papers: RGB-D Grasp Detection via Depth Guided Learning wi…

200 papers

Recently, robotic grasp detection (GD) and object detection (OD) with reasoning have been investigated using deep neural networks (DNNs). There have been works to combine these multi-tasks using separate networks so that robots can deal…

Computer Vision and Pattern Recognition · Computer Science 2019-09-17 Dongwon Park , Yonghyeok Seo , Dongju Shin , Jaesik Choi , Se Young Chun

Accurate depth estimation remains an open problem for robotic manipulation; even state of the art techniques including structured light and LiDAR sensors fail on reflective or transparent surfaces. We address this problem by training a…

Computer Vision and Pattern Recognition · Computer Science 2020-06-17 Ben Goodrich , Alex Kuefler , William D. Richards

Robust semantic perception for autonomous vehicles relies on effectively combining multiple sensors with complementary strengths and weaknesses. State-of-the-art sensor fusion approaches to semantic perception often treat sensor data…

Computer Vision and Pattern Recognition · Computer Science 2026-01-27 Tim Broedermannn , Christos Sakaridis , Luigi Piccinelli , Wim Abbeloos , Luc Van Gool

This work provides an architecture that incorporates depth and tactile information to create rich and accurate 3D models useful for robotic manipulation tasks. This is accomplished through the use of a 3D convolutional neural network (CNN).…

Robotics · Computer Science 2023-02-13 David Watkins , Jacob Varley , Peter Allen

Different from RGB videos, depth data in RGB-D videos provide key complementary information for tristimulus visual data which potentially could achieve accuracy improvement for action recognition. However, most of the existing action…

Computer Vision and Pattern Recognition · Computer Science 2018-11-27 Haokui Zhang , Ying Li , Peng Wang , Yu Liu , Chunhua Shen

Pose-based action recognition has drawn considerable attention recently. Existing methods exploit the joint positions to extract the body-part features from the activation map of the convolutional networks to assist human action…

Computer Vision and Pattern Recognition · Computer Science 2019-12-02 Lei Shi , Yifan Zhang , Jian Cheng , Hanqing Lu

RGB video object tracking is a fundamental task in computer vision. Its effectiveness can be improved using depth information, particularly for handling motion-blurred target. However, depth information is often missing in commonly used…

Computer Vision and Pattern Recognition · Computer Science 2024-10-29 Yu Liu , Arif Mahmood , Muhammad Haris Khan

Graph convolution networks (GCN) have been widely used in skeleton-based action recognition. We note that existing GCN-based approaches primarily rely on prescribed graphical structures (ie., a manually defined topology of skeleton joints),…

Computer Vision and Pattern Recognition · Computer Science 2022-10-13 Haodong Duan , Jiaqi Wang , Kai Chen , Dahua Lin

We propose a new 6-DoF grasp pose synthesis approach from 2D/2.5D input based on keypoints. Keypoint-based grasp detector from image input has demonstrated promising results in the previous study, where the additional visual information…

Robotics · Computer Science 2023-05-02 Yiye Chen , Ruinian Xu , Yunzhi Lin , Hongyi Chen , Patricio A. Vela

Recognizing the category of the object and using the features of the object itself to predict grasp configuration is of great significance to improve the accuracy of the grasp detection model and expand its application. Researchers have…

Robotics · Computer Science 2022-03-03 Mingshuai Dong , Shimin Wei , Jianqin Yin , Xiuli Yu

This paper focuses on the problem of learning 6-DOF grasping with a parallel jaw gripper in simulation. We propose the notion of a geometry-aware representation in grasping based on the assumption that knowledge of 3D geometry is at the…

Real-world robotic grasping can be done robustly if a complete 3D Point Cloud Data (PCD) of an object is available. However, in practice, PCDs are often incomplete when objects are viewed from few and sparse viewpoints before the grasping…

In the realm of future home-assistant robots, 3D articulated object manipulation is essential for enabling robots to interact with their environment. Many existing studies make use of 3D point clouds as the primary input for manipulation…

Robotics · Computer Science 2023-10-16 Xiaoqi Li , Yanzi Wang , Yan Shen , Ponomarenko Iaroslav , Haoran Lu , Qianxu Wang , Boshi An , Jiaming Liu , Hao Dong

Multi-modal RGB and Depth (RGBD) data are predominant in many domains such as robotics, autonomous driving and remote sensing. The combination of these multi-modal data enhances environmental perception by providing 3D spatial context,…

Computer Vision and Pattern Recognition · Computer Science 2025-06-02 Roger Ferrod , Cássio F. Dantas , Luigi Di Caro , Dino Ienco

We propose NormalGAN, a fast adversarial learning-based method to reconstruct the complete and detailed 3D human from a single RGB-D image. Given a single front-view RGB-D image, NormalGAN performs two steps: front-view RGB-D rectification…

Computer Vision and Pattern Recognition · Computer Science 2020-07-31 Lizhen Wang , Xiaochen Zhao , Tao Yu , Songtao Wang , Yebin Liu

Multimodal perception systems for robotics and embodied AI often assume reliable RGB-D sensing, but in practice, depth is frequently missing, noisy, or corrupted. We thus present GeomPrompt, a lightweight cross-modal adaptation module that…

Computer Vision and Pattern Recognition · Computer Science 2026-04-14 Krishna Jaganathan , Patricio Vela

Camouflaged Object Detection (COD) aims to segment objects that are highly integrated with the background in terms of color, texture, and structure, making it a highly challenging task in computer vision. Although existing methods introduce…

Computer Vision and Pattern Recognition · Computer Science 2026-03-26 Min Zhang

Graph Convolution Networks (GCNs) are becoming more and more popular for learning node representations on graphs. Though there exist various developments on sampling and aggregation to accelerate the training process and improve the…

Machine Learning · Computer Science 2020-10-30 Xu Zou , Qiuye Jia , Jianwei Zhang , Chang Zhou , Hongxia Yang , Jie Tang

In multi-robot collaborative area search, a key challenge is to dynamically balance the two objectives of exploring unknown areas and covering specific targets to be rescued. Existing methods are often constrained by homogeneous graph…

Robotics · Computer Science 2026-01-08 Lina Zhu , Jiyu Cheng , Yuehu Liu , Wei Zhang

In this paper, we present Fusion-GCN, an approach for multimodal action recognition using Graph Convolutional Networks (GCNs). Action recognition methods based around GCNs recently yielded state-of-the-art performance for skeleton-based…

Computer Vision and Pattern Recognition · Computer Science 2021-09-28 Michael Duhme , Raphael Memmesheimer , Dietrich Paulus
‹ Prev 1 8 9 10 Next ›