中文
相关论文

相关论文: CLASP: General-Purpose Clothes Manipulation with S…

200 篇论文

General-purpose robotic manipulation, including reach and grasp, is essential for deployment into households and workspaces involving diverse and evolving tasks. Recent advances propose using large pre-trained models, such as Large Language…

机器人学 · 计算机科学 2025-07-16 Huiyi Wang , Fahim Shahriar , Alireza Azimi , Gautham Vasan , Rupam Mahmood , Colin Bellinger

Unsupervised 3D representation learning reduces the burden of labeling multimodal 3D data for fusion perception tasks. Among different pre-training paradigms, differentiable-rendering-based methods have shown most promise. However, existing…

计算机视觉与模式识别 · 计算机科学 2026-03-02 Runjian Chen , Hang Zhang , Avinash Ravichandran , Hyoungseob Park , Wenqi Shao , Alex Wong , Ping Luo

Cloth in the real world is often crumpled, self-occluded, or folded in on itself such that key regions, such as corners, are not directly graspable, making manipulation difficult. We propose a system that leverages visual and tactile…

机器人学 · 计算机科学 2022-12-13 Neha Sunil , Shaoxiong Wang , Yu She , Edward Adelson , Alberto Rodriguez

Garments are important to humans. A visual system that can estimate and track the complete garment pose can be useful for many downstream tasks and real-world applications. In this work, we present a complete package to address the…

计算机视觉与模式识别 · 计算机科学 2025-04-16 Han Xue , Wenqiang Xu , Jieyi Zhang , Tutian Tang , Yutong Li , Wenxin Du , Ruolin Ye , Cewu Lu

Trained on a vast amount of data, Large Language models (LLMs) have achieved unprecedented success and generalization in modeling fairly complex textual inputs in the abstract space, making them powerful tools for zero-shot learning. Such…

计算机视觉与模式识别 · 计算机科学 2023-05-19 Shervin Ardeshir

Multimodal Large Language Models (MLLMs) suffer from substantial computational overhead due to the high redundancy in visual token sequences. Existing approaches typically address this issue using single-layer Vision Transformer (ViT)…

计算机视觉与模式识别 · 计算机科学 2026-04-15 Yunkai Dang , Yizhu Jiang , Yifan Jiang , Qi Fan , Yinghuan Shi , Wenbin Li , Yang Gao

We consider the task of semantic robotic grasping, in which a robot picks up an object of a user-specified class using only monocular images. Inspired by the two-stream hypothesis of visual reasoning, we present a semantic grasping…

机器人学 · 计算机科学 2017-11-10 Eric Jang , Sudheendra Vijayanarasimhan , Peter Pastor , Julian Ibarz , Sergey Levine

Robotic manipulation tasks, such as wiping with a soft sponge, require control from multiple rich sensory modalities. Human-robot interaction, aimed at teaching robots, is difficult in this setting as there is potential for mismatch between…

机器人学 · 计算机科学 2021-03-29 Yordan Hristov , Subramanian Ramamoorthy

Robots can generalize manipulation skills between different scenarios by adapting to the features of the objects being manipulated. Selecting the set of relevant features for generalizing skills has usually been performed manually by a…

机器人学 · 计算机科学 2016-05-17 Oliver Kroemer , Gaurav S. Sukhatme

In the rapidly evolving field of online fashion shopping, the need for more personalized and interactive image retrieval systems has become paramount. Existing methods often struggle with precisely manipulating specific garment attributes…

计算机视觉与模式识别 · 计算机科学 2024-09-17 Vittorio Casula , Lorenzo Berlincioni , Luca Cultrera , Federico Becattini , Chiara Pero , Carmen Bisogni , Marco Bertini , Alberto Del Bimbo

Cluttered garments manipulation poses significant challenges due to the complex, deformable nature of garments and intricate garment relations. Unlike single-garment manipulation, cluttered scenarios require managing complex garment…

机器人学 · 计算机科学 2025-05-19 Ruihai Wu , Ziyu Zhu , Yuran Wang , Yue Chen , Jiarui Wang , Hao Dong

Robot-assisted dressing offers an opportunity to benefit the lives of many people with disabilities, such as some older adults. However, robots currently lack common sense about the physical implications of their actions on people. The…

机器人学 · 计算机科学 2019-05-28 Zackory Erickson , Henry M. Clever , Greg Turk , C. Karen Liu , Charles C. Kemp

Computer vision tasks typically involve describing what is present in an image (e.g. classification, detection, segmentation, and captioning). We study a visual common sense task that requires understanding what is not present.…

计算机视觉与模式识别 · 计算机科学 2024-01-17 Ram Ramrakhya , Aniruddha Kembhavi , Dhruv Batra , Zsolt Kira , Kuo-Hao Zeng , Luca Weihs

Loop closure, as one of the crucial components in SLAM, plays an essential role in correcting the accumulated errors. Traditional appearance-based methods, such as bag-of-words models, are often limited by local 2D features and the volume…

计算机视觉与模式识别 · 计算机科学 2023-11-10 Zhenzhong Cao

From early Movement Primitive (MP) techniques to modern Vision-Language Models (VLMs), autonomous manipulation has remained a pivotal topic in robotics. As two extremes, VLM-based methods emphasize zero-shot and adaptive manipulation but…

机器人学 · 计算机科学 2025-03-05 Junjie Zhu , Huayu Liu , Jin Wang , Bangrong Wen , Kaixiang Huang , Xiaofei Li , Haiyun Zhan , Guodong Lu

Despite significant progress in robotic systems for operation within human-centric environments, existing models still heavily rely on explicit human commands to identify and manipulate specific objects. This limits their effectiveness in…

机器人学 · 计算机科学 2024-10-16 Shiyu Jin , Jinxuan Xu , Yutian Lei , Liangjun Zhang

The primary goal of artificial intelligence is to mimic humans. Therefore, to advance toward this goal, the AI community attempts to imitate qualities/skills possessed by humans and imbibes them into machines with the help of…

计算机视觉与模式识别 · 计算机科学 2022-11-29 Binoy Saha , Sukhendu Das

Bimanual manipulation is imperative yet challenging for robots to execute complex tasks, requiring coordinated collaboration between two arms. However, existing methods for bimanual manipulation often rely on costly data collection and…

机器人学 · 计算机科学 2026-02-11 Jinxian Zhou , Ruihai Wu , Yiwei Liu , Yiwen Hou , Xunzhe Zhou , Checheng Yu , Licheng Zhong , Lin Shao

Due to the high dimensionality of object states, a garment flattening pipeline requires recognising the configurations of garments for a robot to produce/select manipulation plans to flatten garments. In this paper, we propose a…

机器人学 · 计算机科学 2022-08-30 Li Duan , Gerardo Aragon-Camarasa

Currently it requires an artist to create 3D human avatars with realistic clothing that can move naturally. Despite progress on 3D scanning and modeling of human bodies, there is still no technology that can easily turn a static scan into…

计算机视觉与模式识别 · 计算机科学 2021-10-20 Qianli Ma , Jinlong Yang , Siyu Tang , Michael J. Black