中文
相关论文

相关论文: IKEA Manuals at Work: 4D Grounding of Assembly Ins…

200 篇论文

Human-designed visual manuals are crucial components in shape assembly activities. They provide step-by-step guidance on how we should move and connect different parts in a convenient and physically-realizable way. While there has been an…

计算机视觉与模式识别 · 计算机科学 2023-02-06 Ruocheng Wang , Yunzhi Zhang , Jiayuan Mao , Ran Zhang , Chin-Yi Cheng , Jiajun Wu

Assembling furniture amounts to solving the discrete-continuous optimization task of selecting the furniture parts to assemble and estimating their connecting poses in a physically realistic manner. The problem is hampered by its…

计算机视觉与模式识别 · 计算机科学 2025-12-02 Jiahao Zhang , Anoop Cherian , Cristian Rodriguez , Weijian Deng , Stephen Gould

The availability of a large labeled dataset is a key requirement for applying deep learning methods to solve various computer vision tasks. In the context of understanding human activities, existing public datasets, while large in size, are…

计算机视觉与模式识别 · 计算机科学 2023-05-18 Yizhak Ben-Shabat , Xin Yu , Fatemeh Sadat Saleh , Dylan Campbell , Cristian Rodriguez-Opazo , Hongdong Li , Stephen Gould

2D assembly diagrams are often abstract and hard to follow, creating a need for intelligent assistants that can monitor progress, detect errors, and provide step-by-step guidance. In mixed reality settings, such systems must recognize…

计算机视觉与模式识别 · 计算机科学 2026-05-28 Zhuchenyang Liu , Yao Zhang , Yu Xiao

Multimodal alignment facilitates the retrieval of instances from one modality when queried using another. In this paper, we consider a novel setting where such an alignment is between (i) instruction steps that are depicted as assembly…

计算机视觉与模式识别 · 计算机科学 2024-12-03 Jiahao Zhang , Anoop Cherian , Yanbin Liu , Yizhak Ben-Shabat , Cristian Rodriguez , Stephen Gould

Humans possess an extraordinary ability to understand and execute complex manipulation tasks by interpreting abstract instruction manuals. For robots, however, this capability remains a substantial challenge, as they cannot interpret…

机器人学 · 计算机科学 2025-10-21 Chenrui Tie , Shengxiang Sun , Jinxuan Zhu , Yiwei Liu , Jingxiang Guo , Yue Hu , Haonan Chen , Junting Chen , Ruihai Wu , Lin Shao

The IKEA Furniture Assembly Environment is one of the first benchmarks for testing and accelerating the automation of complex manipulation tasks. The environment is designed to advance reinforcement learning from simple toy tasks to complex…

机器人学 · 计算机科学 2019-11-19 Youngwoon Lee , Edward S. Hu , Zhengyu Yang , Alex Yin , Joseph J. Lim

In this paper we address the task of recognizing assembly actions as a structure (e.g. a piece of furniture or a toy block tower) is built up from a set of primitive objects. Recognizing the full range of assembly actions requires…

计算机视觉与模式识别 · 计算机科学 2020-12-03 Jonathan D. Jones , Cathryn Cortesa , Amy Shelton , Barbara Landau , Sanjeev Khudanpur , Gregory D. Hager

Autonomous assembly is a crucial capability for robots in many applications. For this task, several problems such as obstacle avoidance, motion planning, and actuator control have been extensively studied in robotics. However, when it comes…

计算机视觉与模式识别 · 计算机科学 2020-03-25 Yichen Li , Kaichun Mo , Lin Shao , Minhyuk Sung , Leonidas Guibas

Assembly hinges on reliably forming connections between parts; yet most robotic approaches plan assembly sequences and part poses while treating connectors as an afterthought. Connections represent the foundational physical constraints of…

Assembling objects from parts requires understanding multimodal instructions, linking them to 3D components, and predicting physically plausible 6-DoF motions for each assembly step. Existing datasets focus on simplified scenarios,…

计算机视觉与模式识别 · 计算机科学 2026-05-14 Danrui Li , Jiahao Zhang , Bernhard Egger , Moitreya Chatterjee , Suhas Lohit , Tim K. Marks , Anoop Cherian

The recent advancements introduced by Large Language Models (LLMs) have transformed how Artificial Intelligence (AI) can support complex, real world tasks, pushing research outside the text boundaries towards multi modal contexts and…

计算机视觉与模式识别 · 计算机科学 2026-03-25 Federico Toschi , Nicolò Brunello , Andrea Sassella , Vincenzo Scotti , Mark James Carman

Utilizing 6DoF(Degrees of Freedom) pose information of an object and its components is critical for object state detection tasks. We present IKEA Object State Dataset, a new dataset that contains IKEA furniture 3D models, RGBD video of the…

计算机视觉与模式识别 · 计算机科学 2021-11-17 Yongzhi Su , Mingxin Liu , Jason Rambach , Antonia Pehrson , Anton Berg , Didier Stricker

The objective of this work is to manipulate visual timelines (e.g. a video) through natural language instructions, making complex timeline editing tasks accessible to non-expert or potentially even disabled users. We call this task…

计算机视觉与模式识别 · 计算机科学 2024-11-20 Alejandro Pardo , Jui-Hsien Wang , Bernard Ghanem , Josef Sivic , Bryan Russell , Fabian Caba Heilbron

Furniture assembly is a crucial yet challenging task for robots, requiring precise dual-arm coordination where one arm manipulates parts while the other provides collaborative support and stabilization. To accomplish this task more…

机器人学 · 计算机科学 2026-01-19 Jiaqi Liang , Yue Chen , Qize Yu , Yan Shen , Haipeng Zhang , Hao Dong , Ruihai Wu

Assistants on assembly tasks show great potential to benefit humans ranging from helping with everyday tasks to interacting in industrial settings. However, evaluation resources in assembly activities are underexplored. To foster system…

There are substantial instructional videos on the Internet, which provide us tutorials for completing various tasks. Existing instructional video datasets only focus on specific steps at the video level, lacking experiential guidelines at…

计算机视觉与模式识别 · 计算机科学 2024-06-27 Jiafeng Liang , Shixin Jiang , Zekun Wang , Haojie Pan , Zerui Chen , Zheng Chu , Ming Liu , Ruiji Fu , Zhongyuan Wang , Bing Qin

A large-scale dataset is essential for learning good features in 3D shape understanding, but there are only a few datasets that can satisfy deep learning training. One of the major reasons is that current tools for annotating per-point…

计算机视觉与模式识别 · 计算机科学 2021-12-28 Sucheng Qian , Liu Liu , Wenqiang Xu , Cewu Lu

Image-guided object assembly represents a burgeoning research topic in computer vision. This paper introduces a novel task: translating multi-view images of a structural 3D model (for example, one constructed with building blocks drawn from…

计算机视觉与模式识别 · 计算机科学 2024-04-26 Hongyu Yan , Yadong Mu

It is desirable to enable robots capable of automatic assembly. Structural understanding of object parts plays a crucial role in this task yet remains relatively unexplored. In this paper, we focus on the setting of furniture assembly from…

机器人学 · 计算机科学 2022-07-07 Rufeng Zhang , Tao Kong , Weihao Wang , Xuan Han , Mingyu You
‹ 上一页 1 2 3 10 下一页 ›