English
Related papers

Related papers: VOFA: Visual Object Goal Pushing with Force-Adapti…

200 papers

Object navigation in open-world environments remains a formidable and pervasive challenge for robotic systems, particularly when it comes to executing long-horizon tasks that require both open-world object detection and high-level task…

Robotics · Computer Science 2025-07-10 Daojie Peng , Jiahang Cao , Qiang Zhang , Jun Ma

In physical human-robot interaction, force feedback has been the most common sensing modality to convey the human intention to the robot. It is widely used in admittance control to allow the human to direct the robot. However, it cannot be…

Robotics · Computer Science 2025-05-16 Gokhan Solak , Gustavo J. G. Lahr , Idil Ozdamar , Arash Ajoudani

Balancing and push-recovery are essential capabilities enabling humanoid robots to solve complex locomotion tasks. In this context, classical control systems tend to be based on simplified physical models and hard-coded strategies. Although…

Humanoid robots deployed in real-world scenarios often need to carry unknown payloads, which introduce significant mismatch and degrade the effectiveness of simulation-to-reality reinforcement learning methods. To address this challenge, we…

Robotics · Computer Science 2026-03-17 Xingyi Wang , Chenyun Zhang , Weiji Xie , Chao Yu , Wei Song , Chenjia Bai , Shiqiang Zhu

Vision-Language-Action (VLA) models have recently advanced robotic manipulation by translating natural-language instructions and visual observations into control actions. However, existing VLAs are primarily trained on successful expert…

Robotics · Computer Science 2026-03-24 Zewei Ye , Weifeng Lu , Minghao Ye , Tao Lin , Shuo Yang , Junchi Yan , Bo Zhao

Catastrophic forgetting is a well-documented challenge in model fine-tuning, particularly when the downstream domain has limited labeled data or differs substantially from the pre-training distribution. Existing parameter-efficient…

Machine Learning · Computer Science 2026-02-03 Peng Wang , Minghao Gu , Qiang Huang

Many robots are not equipped with a manipulator and many objects are not suitable for prehensile manipulation (such as large boxes and cylinders). In these cases, pushing is a simple yet effective non-prehensile skill for robots to interact…

Robotics · Computer Science 2025-11-21 Zili Tang , Ying Zhang , Meng Guo

In complex scenarios where typical pick-and-place techniques are insufficient, often non-prehensile manipulation can ensure that a robot is able to fulfill its task. However, non-prehensile manipulation is challenging due to its…

Robotics · Computer Science 2025-08-04 Nils Dengler , Juan Del Aguila Ferrandis , João Moura , Sethu Vijayakumar , Maren Bennewitz

Moving large objects, such as furniture or appliances, is a critical capability for robots operating in human environments. This task presents unique challenges, including whole-body coordination to avoid collisions and managing the…

Robotics · Computer Science 2025-05-15 Tianyu Li , Joanne Truong , Jimmy Yang , Alexander Clegg , Akshara Rai , Sehoon Ha , Xavier Puig

Non-prehensile manipulation, such as pushing objects to a desired target position, is an important skill for robots to assist humans in everyday situations. However, the task is challenging due to the large variety of objects with different…

Robotics · Computer Science 2024-11-14 Lara Bergmann , David Leins , Robert Haschke , Klaus Neumann

Vision-Language-Action (VLA) models have advanced general-purpose robotic manipulation by leveraging pretrained visual and linguistic representations. However, they struggle with contact-rich tasks that require fine-grained control…

Vision-Language-Action (VLA) models are promising for generalist robot manipulation but remain brittle in out-of-distribution (OOD) settings, especially with limited real-robot data. To resolve the generalization bottleneck, we introduce a…

The complex physical properties of highly deformable materials such as clothes pose significant challenges fanipulation systems. We present a novel visual feedback dictionary-based method for manipulating defoor autonomous robotic mrmable…

Robotics · Computer Science 2019-01-18 Biao Jia , Zhe Hu , Jia Pan , Dinesh Manocha

Whole-body control for humanoids is challenging due to the high-dimensional nature of the problem, coupled with the inherent instability of a bipedal morphology. Learning from visual observations further exacerbates this difficulty. In this…

Machine Learning · Computer Science 2025-05-16 Nicklas Hansen , Jyothir S , Vlad Sobal , Yann LeCun , Xiaolong Wang , Hao Su

This manuscript introduces an object deformability-agnostic framework for co-carrying tasks that are shared between a person and multiple robots. Our approach allows the full control of the co-carrying trajectories by the person while…

Robotics · Computer Science 2023-02-10 Doganay Sirintuna , Idil Ozdamar , Arash Ajoudani

In this work, we aim to learn a unified vision-based policy for multi-fingered robot hands to manipulate a variety of objects in diverse poses. Though prior work has shown benefits of using human videos for policy learning, performance…

Computer Vision and Pattern Recognition · Computer Science 2025-03-04 Zerui Chen , Shizhe Chen , Etienne Arlaud , Ivan Laptev , Cordelia Schmid

Learning to manipulate objects efficiently, particularly those involving sustained contact (e.g., pushing, sliding) and articulated parts (e.g., drawers, doors), presents significant challenges. Traditional methods, such as robot-centric…

Robotics · Computer Science 2025-03-18 Shijie Fang , Wenchang Gao , Shivam Goel , Christopher Thierauf , Matthias Scheutz , Jivko Sinapov

Human-robot collaboration requires the contactless estimation of the physical properties of containers manipulated by a person, for example while pouring content in a cup or moving a food box. Acoustic and visual signals can be used to…

Multimedia · Computer Science 2022-03-07 A. Xompero , Y. L. Pang , T. Patten , A. Prabhakar , B. Calli , A. Cavallaro

PointGoal navigation in indoor environment is a fundamental task for personal robots to navigate to a specified point. Recent studies solved this PointGoal navigation task with near-perfect success rate in photo-realistically simulated…

Computer Vision and Pattern Recognition · Computer Science 2023-04-04 Yijun Cao , Xianshi Zhang , Fuya Luo , Chuan Lin , Yongjie Li

Current vision-language-action (VLA) models, pre-trained on large-scale robotic data, exhibit strong multi-task capabilities and generalize well to variations in visual and language instructions for manipulation. However, their success rate…

Robotics · Computer Science 2025-10-17 Han Zhao , Jiaxuan Zhang , Wenxuan Song , Pengxiang Ding , Donglin Wang