中文
相关论文

相关论文: Learning human-to-robot handovers through 3D scene…

200 篇论文

Vision-based human-to-robot handover is an important and challenging task in human-robot interaction. Recent work has attempted to train robot policies by interacting with dynamic virtual humans in simulated environments, where the policies…

机器人学 · 计算机科学 2025-01-03 Sammy Christen , Lan Feng , Wei Yang , Yu-Wei Chao , Otmar Hilliges , Jie Song

Novel view synthesis for dynamic scenes is still a challenging problem in computer vision and graphics. Recently, Gaussian splatting has emerged as a robust technique to represent static scenes and enable high-quality and real-time novel…

计算机视觉与模式识别 · 计算机科学 2024-04-15 Yi-Hua Huang , Yang-Tian Sun , Ziyi Yang , Xiaoyang Lyu , Yan-Pei Cao , Xiaojuan Qi

Recent advancements in imitation learning have shown promising results in robotic manipulation, driven by the availability of high-quality training data. To improve data collection efficiency, some approaches focus on developing specialized…

Implicit neural representations and 3D Gaussian splatting (3DGS) have shown great potential for scene reconstruction. Recent studies have expanded their applications in autonomous reconstruction through task assignment methods. However,…

机器人学 · 计算机科学 2024-12-04 Jing Zeng , Qi Ye , Tianle Liu , Yang Xu , Jin Li , Jinming Xu , Liang Li , Jiming Chen

Surgical simulation is essential for medical training, enabling practitioners to develop crucial skills in a risk-free environment while improving patient safety and surgical outcomes. However, conventional methods for building simulation…

计算机视觉与模式识别 · 计算机科学 2025-09-23 Zhenya Yang

In complex missions such as search and rescue,robots must make intelligent decisions in unknown environments, relying on their ability to perceive and understand their surroundings. High-quality and real-time reconstruction enhances…

机器人学 · 计算机科学 2024-10-10 Zijun Xu , Rui Jin , Ke Wu , Yi Zhao , Zhiwei Zhang , Jieru Zhao , Fei Gao , Zhongxue Gan , Wenchao Ding

In-hand object reorientation requires precise estimation of the object pose to handle complex task dynamics. While RGB sensing offers rich semantic cues for pose tracking, existing solutions rely on multi-camera setups or costly ray…

机器人学 · 计算机科学 2026-04-14 Arjun Bhardwaj , Maximum Wilder-Smith , Mayank Mittal , Vaishakh Patil , Marco Hutter

An excellent representation is crucial for reinforcement learning (RL) performance, especially in vision-based reinforcement learning tasks. The quality of the environment representation directly influences the achievement of the learning…

计算机视觉与模式识别 · 计算机科学 2024-08-07 Jiaxu Wang , Qiang Zhang , Jingkai Sun , Jiahang Cao , Gang Han , Wen Zhao , Weining Zhang , Yecheng Shao , Yijie Guo , Renjing Xu

Imitation learning is a popular paradigm to teach robots new tasks, but collecting robot demonstrations through teleoperation or kinesthetic teaching is tedious and time-consuming. In contrast, directly demonstrating a task using our human…

机器人学 · 计算机科学 2026-02-16 Nick Heppert , Minh Quang Nguyen , Abhinav Valada

Self-modeling enables robots to build task-agnostic models of their morphology and kinematics based on data that can be automatically collected, with minimal human intervention and prior information, thereby enhancing machine intelligence.…

机器人学 · 计算机科学 2025-11-17 Kejun Hu , Peng Yu , Ning Tan

Training robot policies within a learned world model is trending due to the inefficiency of real-world interactions. The established image-based world models and policies have shown prior success, but lack robust geometric information that…

机器人学 · 计算机科学 2025-09-18 Guanxing Lu , Baoxiong Jia , Puhao Li , Yixin Chen , Ziwei Wang , Yansong Tang , Siyuan Huang

Learning robust and generalizable manipulation skills from demonstrations remains a key challenge in robotics, with broad applications in industrial automation and service robotics. While recent imitation learning methods have achieved…

计算机视觉与模式识别 · 计算机科学 2024-11-18 Yu Ren , Yang Cong , Ronghan Chen , Jiahao Long

Training a deep network policy for robot manipulation is notoriously costly and time consuming as it depends on collecting a significant amount of real world data. To work well in the real world, the policy needs to see many instances of…

机器人学 · 计算机科学 2019-06-24 Xinchen Yan , Mohi Khansari , Jasmine Hsu , Yuanzheng Gong , Yunfei Bai , Sören Pirk , Honglak Lee

Photorealistic 3D scene reconstruction plays an important role in autonomous driving, enabling the generation of novel data from existing datasets to simulate safety-critical scenarios and expand training data without additional acquisition…

计算机视觉与模式识别 · 计算机科学 2025-11-26 Pou-Chun Kung , Xianling Zhang , Katherine A. Skinner , Nikita Jaipuria

Training robotic manipulation policies traditionally requires numerous demonstrations and/or environmental rollouts. While recent Imitation Learning (IL) and Reinforcement Learning (RL) methods have reduced the number of required…

机器人学 · 计算机科学 2025-03-18 Peihong Yu , Amisha Bhaskar , Anukriti Singh , Zahiruddin Mahammad , Pratap Tokekar

We present GASPACHO, a method for generating photorealistic, controllable renderings of human-object interactions from multi-view RGB video. Unlike prior work that reconstructs only the human and treats objects as background, GASPACHO…

计算机视觉与模式识别 · 计算机科学 2025-12-17 Aymen Mir , Arthur Moreau , Helisa Dhamo , Zhensong Zhang , Gerard Pons-Moll , Eduardo Pérez-Pellitero

Reaching-and-grasping is a fundamental skill for robotic manipulation, but existing methods usually train models on a specific gripper and cannot be reused on another gripper. In this paper, we propose a novel method that can learn a…

机器人学 · 计算机科学 2025-02-04 Qijin She , Shishun Zhang , Yunfan Ye , Ruizhen Hu , Kai Xu

3D Gaussian Splatting offers expressive scene reconstruction, modeling a broad range of visual, geometric, and semantic information. However, efficient real-time map reconstruction with data streamed from multiple robots and devices remains…

机器人学 · 计算机科学 2025-06-04 Javier Yu , Timothy Chen , Mac Schwager

The control of a robot for manipulation tasks generally relies on object detection and pose estimation. An attractive alternative is to learn control policies directly from raw input data. However, this approach is time-consuming and…

机器人学 · 计算机科学 2021-08-10 Changjae Oh , Yik Lung Pang , Andrea Cavallaro

Videos of robots interacting with objects encode rich information about the objects' dynamics. However, existing video prediction approaches typically do not explicitly account for the 3D information from videos, such as robot actions and…

机器人学 · 计算机科学 2024-10-25 Mingtong Zhang , Kaifeng Zhang , Yunzhu Li