English
Related papers

Related papers: InterACT: Inter-dependency Aware Action Chunking w…

200 papers

Controllable cooperative humanoid manipulation is a fundamental yet challenging problem for embodied intelligence, due to severe data scarcity, complexities in multi-agent coordination, and limited generalization across objects. In this…

Computer Vision and Pattern Recognition · Computer Science 2026-04-22 Wei Yao , Haohan Ma , Hongwen Zhang , Yunlian Sun , Liangjun Xing , Zhile Yang , Yuanjun Guo , Yebin Liu , Jinhui Tang

In-hand manipulation is challenging for a multi-finger robotic hand due to its high degrees of freedom and the complex interaction with the object. To enable in-hand manipulation, existing deep reinforcement learning based approaches mainly…

Robotics · Computer Science 2023-07-12 Lingfeng Tao , Jiucai Zhang , Michael Bowman , Xiaoli Zhang

Human Activity Recognition (HAR) using wearable devices such as smart watches embedded with Inertial Measurement Unit (IMU) sensors has various applications relevant to our daily life, such as workout tracking and health monitoring. In this…

Computer Vision and Pattern Recognition · Computer Science 2021-12-22 Wenjin Tao , Haodong Chen , Md Moniruzzaman , Ming C. Leu , Zhaozheng Yi , Ruwen Qin

Bidirectional transformers are the foundation of many sequence modeling tasks across natural, biological, and chemical language domains, but they are permutation-invariant without explicit positional embeddings. In contrast, unidirectional…

Quantitative Methods · Quantitative Biology 2026-04-22 Logan Hallee , Jason P. Gleghorn

Recent advances in multimodal vision-language-action (VLA) models have revolutionized traditional robot learning, enabling systems to interpret vision, language, and action in unified frameworks for complex task planning. However, mastering…

Robotics · Computer Science 2025-06-12 Hongjun Wu , Heng Zhang , Pengsong Zhang , Jin Wang , Cong Wang

In open-ended continuous environments, robots need to learn multiple parameterised control tasks in hierarchical reinforcement learning. We hypothesise that the most complex tasks can be learned more easily by transferring knowledge from…

Artificial Intelligence · Computer Science 2021-02-22 Nicolas Duminy , Sao Mai Nguyen , Junshuai Zhu , Dominique Duhaut , Jerome Kerdreux

Most successes in robotic manipulation have been restricted to single-arm robots, which limits the range of solvable tasks to pick-and-place, insertion, and objects rearrangement. In contrast, dual and multi arm robot platforms unlock a…

Robotics · Computer Science 2022-03-17 Satoshi Kataoka , Seyed Kamyar Seyed Ghasemipour , Daniel Freeman , Igor Mordatch

Symmetric bi-manual manipulation is an essential skill in on-orbit operations due to its potent load capacity. Previous works have applied compliant control to maintain the stability of manipulations. However, traditional methods have…

Robotics · Computer Science 2024-07-22 Yuxue Cao , Wenbo Zhao , Shengjie Wang , Xiang Zheng , Wenke Ma , Zhaolei Wang , Tao Zhang

Manipulation and locomotion are closely related problems that are often studied in isolation. In this work, we study the problem of coordinating multiple mobile agents to exhibit manipulation behaviors using a reinforcement learning (RL)…

Robotics · Computer Science 2019-10-09 Ofir Nachum , Michael Ahn , Hugo Ponte , Shixiang Gu , Vikash Kumar

People can learn a wide range of tasks from their own experience, but can also learn from observing other creatures. This can accelerate acquisition of new skills even when the observed agent differs substantially from the learning agent in…

Artificial Intelligence · Computer Science 2017-03-09 Abhishek Gupta , Coline Devin , YuXuan Liu , Pieter Abbeel , Sergey Levine

Autonomous manipulation in robot arms is a complex and evolving field of study in robotics. This paper introduces an innovative approach to this challenge by focusing on imitation learning (IL). Unlike traditional imitation methods, our…

Robotics · Computer Science 2024-02-06 Masato Kobayashi , Thanpimon Buamanee , Yuki Uranishi , Haruo Takemura

Human Object Interaction (HOI) detection is a challenging task that requires to distinguish the interaction between a human-object pair. Attention based relation parsing is a popular and effective strategy utilized in HOI. However, current…

Computer Vision and Pattern Recognition · Computer Science 2022-07-19 Jingjia Huang , Baixiang Yang

Recent video generation research has focused heavily on isolated actions, leaving interactive motions-such as hand-face interactions-largely unexamined. These interactions are essential for emerging biometric authentication systems, which…

Computer Vision and Pattern Recognition · Computer Science 2025-08-19 Yukang Lin , Yan Hong , Zunnan Xu , Xindi Li , Chao Xu , Chuanbiao Song , Ronghui Li , Haoxing Chen , Jun Lan , Huijia Zhu , Weiqiang Wang , Jianfu Zhang , Xiu Li

Mobile manipulators are designed to perform complex sequences of navigation and manipulation tasks in human-centered environments. While recent optimization-based methods such as Hierarchical Task Model Predictive Control (HTMPC) enable…

Robotics · Computer Science 2026-05-29 Francesco D'Orazio , Sepehr Samavi , Xintong Du , Siqi Zhou , Giuseppe Oriolo , Angela P. Schoellig

This study proposes an imitation learning method based on force and position information. Force information is required for precise object manipulation but is difficult to obtain because the acting and reaction forces cannnot be separated.…

Robotics · Computer Science 2018-11-29 Tsuyoshi Adachi , Kazuki Fujimoto , Sho Sakaino , Toshiaki Tsuji

In human-robot collaboration, shared control presents an opportunity to teleoperate robotic manipulation to improve the efficiency of manufacturing and assembly processes. Robots are expected to assist in executing the user's intentions. To…

Robotics · Computer Science 2024-04-01 Mingyu Cai , Karankumar Patel , Soshi Iba , Songpo Li

Transformer-based methods have shown impressive performance in low-level vision tasks, such as image super-resolution. However, we find that these networks can only utilize a limited spatial range of input information through attribution…

Image and Video Processing · Electrical Eng. & Systems 2023-03-21 Xiangyu Chen , Xintao Wang , Jiantao Zhou , Yu Qiao , Chao Dong

There has recently been significant interest in training reinforcement learning (RL) agents in vision-based environments. This poses many challenges, such as high dimensionality and the potential for observational overfitting through…

This paper proposes a novel method for understanding daily hand-object manipulation by developing computer vision-based techniques. Specifically, we focus on recognizing hand grasp types, object attributes and manipulation actions within an…

Computer Vision and Pattern Recognition · Computer Science 2018-07-24 Minjie Cai , Kris Kitani , Yoichi Sato

Various works have aimed at combining the inference efficiency of recurrent models and training parallelism of multi-head attention for sequence modeling. However, most of these works focus on tasks with fixed-dimension observation spaces,…

Machine Learning · Computer Science 2024-10-14 Bryce Ferenczi , Michael Burke , Tom Drummond