中文
相关论文

相关论文: MRG: A Multi-Robot Manufacturing Digital Scene Gen…

200 篇论文

On robotics computer vision tasks, generating and annotating large amounts of data from real-world for the use of deep learning-based approaches is often difficult or even impossible. A common strategy for solving this problem is to apply…

计算机视觉与模式识别 · 计算机科学 2023-01-13 Chengzhi Wu , Xuelei Bi , Julius Pfrommer , Alexander Cebulla , Simon Mangold , Jürgen Beyerer

Object recognition and object pose estimation in robotic grasping continue to be significant challenges, since building a labelled dataset can be time consuming and financially costly in terms of data collection and annotation. In this…

计算机视觉与模式识别 · 计算机科学 2024-01-25 Dongmyoung Lee , Wei Chen , Nicolas Rojas

Due to their complex spatial structure and diverse geometric features, achieving high-precision and robust point cloud registration for complex Die Castings has been a significant challenge in the die-casting industry. Existing point cloud…

计算机视觉与模式识别 · 计算机科学 2024-03-18 Yu Du , Yu Song , Ce Guo , Xiaojing Tian , Dong Liu , Ming Cong

Accurate temporal prediction is the bridge between comprehensive scene understanding and embodied artificial intelligence. However, predicting multiple fine-grained states of a scene at multiple temporal scales is difficult for…

计算机视觉与模式识别 · 计算机科学 2026-01-27 Zhitao Zeng , Guojian Yuan , Junyuan Mao , Yuxuan Wang , Xiaoshuang Jia , Yueming Jin

In recent years, point cloud generation has gained significant attention in 3D generative modeling. Among existing approaches, point-based methods directly generate point clouds without relying on other representations such as latent…

计算机视觉与模式识别 · 计算机科学 2025-11-26 Petr Molodyk , Jaemoo Choi , David W. Romero , Ming-Yu Liu , Yongxin Chen

While text-to-video diffusion models have advanced significantly, creating coherent long-form content remains unreliable due to stochastic sampling artifacts. This necessitates generating multiple candidates, yet verifying them creates a…

计算机视觉与模式识别 · 计算机科学 2026-04-09 Daewon Yoon , Hyeongseok Lee , Wonsik Shin , Sangyu Han , Nojun Kwak

Representing the environment is a central challenge in robotics, and is essential for effective decision-making. Traditionally, before capturing images with a manipulator-mounted camera, users need to calibrate the camera using a specific…

机器人学 · 计算机科学 2024-04-19 Weiming Zhi , Haozhan Tang , Tianyi Zhang , Matthew Johnson-Roberson

This paper proposes a mode multigrid (MMG) method, and applies it to accelerate the convergence of the steady state flow on unstructured grids. The dynamic mode decomposition (DMD) technique is used to analyze the convergence process of…

计算物理 · 物理学 2018-02-27 Yilang Liu , Weiwei Zhang , Jiaqing Kou

Multiview point cloud registration serves as a cornerstone of various computer vision tasks. Previous approaches typically adhere to a global paradigm, where a pose graph is initially constructed followed by motion synchronization to…

计算机视觉与模式识别 · 计算机科学 2024-07-11 Shiqi Li , Jihua Zhu , Yifan Xie , Mingchen Zhu

Accurate three-dimensional perception is a fundamental task in several computer vision applications. Recently, commercial RGB-depth (RGB-D) cameras have been widely adopted as single-view depth-sensing devices owing to their efficient…

计算机视觉与模式识别 · 计算机科学 2022-07-27 Jiwan Kim , Minchang Kim , Yeong-Gil Shin , Minyoung Chung

Point cloud registration aims to provide estimated transformations to align point clouds, which plays a crucial role in pose estimation of various navigation systems, such as surgical guidance systems and autonomous vehicles. Despite the…

计算机视觉与模式识别 · 计算机科学 2024-10-02 Geng Li , Haozhi Cao , Mingyang Liu , Shenghai Yuan , Jianfei Yang

Operating rooms (ORs) are complex, high-stakes environments requiring precise understanding of interactions among medical staff, tools, and equipment for enhancing surgical assistance, situational awareness, and patient safety. Current…

计算机视觉与模式识别 · 计算机科学 2025-03-05 Ege Özsoy , Chantal Pellegrini , Tobias Czempiel , Felix Tristram , Kun Yuan , David Bani-Harouni , Ulrich Eck , Benjamin Busam , Matthias Keicher , Nassir Navab

Multi-instance point cloud registration aims to estimate the pose of all instances of a model point cloud in the whole scene. Existing methods all adopt the strategy of first obtaining the global correspondence and then clustering to obtain…

计算机视觉与模式识别 · 计算机科学 2025-01-07 Liyuan Zhang , Le Hui , Qi Liu , Bo Li , Yuchao Dai

Accurate grasping is the key to several robotic tasks including assembly and household robotics. Executing a successful grasp in a cluttered environment requires multiple levels of scene understanding: First, the robot needs to analyze the…

机器人学 · 计算机科学 2024-05-13 René Zurbrügg , Yifan Liu , Francis Engelmann , Suryansh Kumar , Marco Hutter , Vaishakh Patil , Fisher Yu

Point set registration is an essential step in many computer vision applications, such as 3D reconstruction and SLAM. Although there exist many registration algorithms for different purposes, however, this topic is still challenging due to…

计算机视觉与模式识别 · 计算机科学 2022-11-22 Jin Zhang , Mingyang Zhao , Xin Jiang , Dong-Ming Yan

Object-Centric Motion Generation (OCMG) plays a key role in a variety of industrial applications$\unicode{x2014}$such as robotic spray painting and welding$\unicode{x2014}$requiring efficient, scalable, and generalizable algorithms to plan…

机器人学 · 计算机科学 2025-02-27 Gabriele Tiboni , Raffaello Camoriano , Tatiana Tommasi

This paper concerns the research problem of point cloud registration to find the rigid transformation to optimally align the source point set with the target one. Learning robust point cloud registration models with deep neural networks has…

计算机视觉与模式识别 · 计算机科学 2024-02-23 Yu Hao , Yi Fang

Generalising vision-based manipulation policies to novel environments remains a challenging area with limited exploration. Current practices involve collecting data in one location, training imitation learning or reinforcement learning…

机器人学 · 计算机科学 2024-09-10 Eugene Teoh , Sumit Patidar , Xiao Ma , Stephen James

Robotic manipulation systems benefit from complementary sensing modalities, where each provides unique environmental information. Point clouds capture detailed geometric structure, while RGB images provide rich semantic context. Current…

Diffusion models have recently been successfully applied to a wide range of robotics applications for learning complex multi-modal behaviors from data. However, prior works have mostly been confined to single-robot and small-scale…

机器人学 · 计算机科学 2025-05-08 Yorai Shaoul , Itamar Mishani , Shivam Vats , Jiaoyang Li , Maxim Likhachev