中文
相关论文

相关论文: KOVIS: Keypoint-based Visual Servoing with Zero-Sh…

200 篇论文

This work introduces Robots Imitating Generated Videos (RIGVid), a system that enables robots to perform complex manipulation tasks--such as pouring, wiping, and mixing--purely by imitating AI-generated videos, without requiring any…

机器人学 · 计算机科学 2026-05-14 Shivansh Patel , Shraddhaa Mohan , Hanlin Mai , Unnat Jain , Svetlana Lazebnik , Yunzhu Li

We present a system for learning a challenging dexterous manipulation task involving moving a cube to an arbitrary 6-DoF pose with only 3-fingers trained with NVIDIA's IsaacGym simulator. We show empirical benefits, both in simulation and…

The machine vision systems have been playing a significant role in visual monitoring systems. With the help of stereovision and machine learning, it will be able to mimic human-like visual system and behaviour towards the environment. In…

计算机视觉与模式识别 · 计算机科学 2024-07-01 Mohamed Fazil M. S. , Arockia Selvakumar A. , Daniel Schilberg

Indirect simultaneous positioning (ISP), where internal tissue points are placed at desired locations indirectly through the manipulation of boundary points, is a type of subtask frequently performed in robotic surgeries. Although…

机器人学 · 计算机科学 2023-06-27 Yafei Ou , Mahdi Tavakoli

3D Move To See (3DMTS) is a mutli-perspective visual servoing method for unstructured and occluded environments, like that encountered in robotic crop harvesting. This paper presents a deep learning method, Deep-3DMTS for creating a…

机器人学 · 计算机科学 2019-08-07 Paul Zapotezny-Anderson , Chris Lehnert

This paper considers the final approach phase of visual-closed-loop grasping where the RGB-D camera is no longer able to provide valid depth information. Many current robotic grasping controllers are not closed-loop and therefore fail for…

机器人学 · 计算机科学 2020-03-02 Jesse Haviland , Feras Dayoub , Peter Corke

Manipulation of deformable objects, such as ropes and cloth, is an important but challenging problem in robotics. We present a learning-based system where a robot takes as input a sequence of images of a human manipulating a rope from an…

计算机视觉与模式识别 · 计算机科学 2017-03-07 Ashvin Nair , Dian Chen , Pulkit Agrawal , Phillip Isola , Pieter Abbeel , Jitendra Malik , Sergey Levine

Purpose: A profound education of novice surgeons is crucial to ensure that surgical interventions are effective and safe. One important aspect is the teaching of technical skills for minimally invasive or robot-assisted procedures. This…

计算机视觉与模式识别 · 计算机科学 2019-09-05 Isabel Funke , Sören Torge Mees , Jürgen Weitz , Stefanie Speidel

Multi-View Stereo (MVS) is a core task in 3D computer vision. With the surge of novel deep learning methods, learned MVS has surpassed the accuracy of classical approaches, but still relies on building a memory intensive dense cost volume.…

计算机视觉与模式识别 · 计算机科学 2022-06-16 Radu Alexandru Rosu , Sven Behnke

Recent advances in robotic learning in simulation have shown impressive results in accelerating learning complex manipulation skills. However, the sim-to-real gap, caused by discrepancies between simulation and reality, poses significant…

机器人学 · 计算机科学 2025-03-25 Jacinto Colan , Keisuke Sugita , Ana Davila , Yutaro Yamada , Yasuhisa Hasegawa

State-of-the-art computer vision systems are trained to predict a fixed set of predetermined object categories. This restricted form of supervision limits their generality and usability since additional labeled data is needed to specify any…

This work proposes a novel method for object co-segmentation, i.e. pixel-level localization of a common object in a set of images, that uses no pixel-level supervision for training. Two pre-trained Vision Transformer (ViT) models are…

计算机视觉与模式识别 · 计算机科学 2024-10-18 Nikolaos-Antonios Ypsilantis , Ondřej Chum

When developing control laws for robotic systems, the principle factor when examining their performance is choosing inputs that allow smooth tracking to a reference input. In the context of robotic manipulation, this involves translating an…

机器人学 · 计算机科学 2026-04-02 Ethan Canzini , Simon Pope , Ashutosh Tiwari

Limited power and computational resources, absence of high-end sensor equipment and GPS-denied environments are challenges faced by autonomous micro areal vehicles (MAVs). We address these challenges in the context of autonomous navigation…

机器人学 · 计算机科学 2020-09-10 Max Christl

Learning robot manipulation policies from raw, real-world image data requires a large number of robot-action trials in the physical environment. Although training using simulations offers a cost-effective alternative, the visual domain gap…

机器人学 · 计算机科学 2025-07-14 Yuekun Wu , Yik Lung Pang , Andrea Cavallaro , Changjae Oh

Grasping is a fundamental task in robot-assisted surgery (RAS), and automating it can reduce surgeon workload while enhancing efficiency, safety, and consistency beyond teleoperated systems. Most prior approaches rely on explicit object…

机器人学 · 计算机科学 2025-08-18 Hongbin Lin , Bin Li , Kwok Wai Samuel Au

Deformable object manipulation is a classical and challenging research area in robotics. Compared with rigid object manipulation, this problem is more complex due to the deformation properties including elastic, plastic, and elastoplastic…

机器人学 · 计算机科学 2024-05-14 Jianhua Shan , Yuhao Sun , Shixin Zhang , Fuchun Sun , Zixi Chen , Zirong Shen , Cesare Stefanini , Yiyong Yang , Shan Luo , Bin Fang

This paper describes a stereo image-based visual servoing system for trajectory tracking by a non-holonomic robot without externally derived pose information nor a known visual map of the environment. It is called trajectory servoing. The…

机器人学 · 计算机科学 2022-06-15 Shiyu Feng , Zixuan Wu , Yipu Zhao , Patricio A. Vela

Tasks such as autonomous navigation, 3D reconstruction, and object recognition near the water surfaces are crucial in marine robotics applications. However, challenges arise due to dynamic disturbances, e.g., light reflections and…

计算机视觉与模式识别 · 计算机科学 2025-04-16 Jiayi Wu , Xiaomin Lin , Shahriar Negahdaripour , Cornelia Fermüller , Yiannis Aloimonos

Visual robotic manipulation research and applications often use multiple cameras, or views, to better perceive the world. How else can we utilize the richness of multi-view data? In this paper, we investigate how to learn good…

机器人学 · 计算机科学 2023-06-01 Younggyo Seo , Junsu Kim , Stephen James , Kimin Lee , Jinwoo Shin , Pieter Abbeel
‹ 上一页 1 8 9 10 下一页 ›