English
Related papers

Related papers: DYNAMO: Dependency-Aware Deep Learning Framework f…

200 papers

Synthesizing text-driven 3D human motion within realistic scenes requires learning both semantic intent ("walk to the couch") and physical feasibility (e.g., avoiding collisions). Current methods use generative frameworks that…

Computer Vision and Pattern Recognition · Computer Science 2026-02-25 Anindita Ghosh , Vladislav Golyanik , Taku Komura , Philipp Slusallek , Christian Theobalt , Rishabh Dabral

Prior work for articulated 3D shape reconstruction often relies on specialized sensors (e.g., synchronized multi-camera systems), or pre-built 3D deformable models (e.g., SMAL or SMPL). Such methods are not able to scale to diverse sets of…

Computer Vision and Pattern Recognition · Computer Science 2023-04-04 Gengshan Yang , Minh Vo , Natalia Neverova , Deva Ramanan , Andrea Vedaldi , Hanbyul Joo

In this paper, we propose a novel framework, Combo, for harmonious co-speech holistic 3D human motion generation and efficient customizable adaption. In particular, we identify that one fundamental challenge as the…

Computer Vision and Pattern Recognition · Computer Science 2025-09-22 Chao Xu , Mingze Sun , Zhi-Qi Cheng , Fei Wang , Yang Liu , Baigui Sun , Ruqi Huang , Alexander Hauptmann

Robots can rapidly acquire new skills from demonstrations. However, during generalisation of skills or transitioning across fundamentally different skills, it is unclear whether the robot has the necessary knowledge to perform the task.…

Machine Learning · Statistics 2018-08-08 Nutan Chen , Alexej Klushyn , Alexandros Paraschos , Djalel Benbouzid , Patrick van der Smagt

This paper presents DINO-RotateMatch, a deep-learning framework designed to address the chal lenges of image matching in large-scale 3D reconstruction from unstructured Internet images. The method integrates a dataset-adaptive image pairing…

Computer Vision and Pattern Recognition · Computer Science 2025-12-04 Kaichen Zhang , Tianxiang Sheng , Xuanming Shi

Motion prediction is a key factor towards the full deployment of autonomous vehicles. It is fundamental in order to assure safety while navigating through highly interactive complex scenarios. In this work, the framework IAMP (Interaction-…

Robotics · Computer Science 2023-04-25 Vinicius Trentin , Chenxu Ma , Jorge Villagra , Zaid Al-Ars

Recent advancements in large-scale generative models have significantly improved the quality and diversity of 3D shape generation. However, most existing methods focus primarily on generating static 3D models, overlooking the potentially…

Computer Vision and Pattern Recognition · Computer Science 2025-03-27 Mingze Sun , Shiwei Mao , Keyi Chen , Yurun Chen , Shunlin Lu , Jingbo Wang , Junting Dong , Ruqi Huang

Contemporary approaches to solving various problems that require analyzing three-dimensional (3D) meshes and point clouds have adopted the use of deep learning algorithms that directly process 3D data such as point coordinates, normal…

Computer Vision and Pattern Recognition · Computer Science 2024-07-12 Stefan Novaković , Vladimir Risojević

From the complex motions of robots to the oxygen binding of hemoglobin, the function of many mechanical systems depends on large, coordinated movements of their components. Such movements arise from a network of physical interactions in the…

Soft Condensed Matter · Physics 2019-06-21 Jason Z. Kim , Zhixin Lu , Danielle S. Bassett

Learning from Demonstration (LfD) has shown to provide robots with fundamental motion skills for a variety of domains. Various branches of LfD research (e.g., learned dynamical systems and movement primitives) can generally be classified…

Robotics · Computer Science 2025-11-20 Alex Cuellar , Christopher K Fourie , Julie A Shah

Robots operating in domestic environments generally need to interact with articulated objects, such as doors, cabinets, dishwashers or fridges. In this work, we present a novel, probabilistic framework for modeling articulated objects as…

Robotics · Computer Science 2014-06-02 Jürgen Sturm , Cyrill Stachniss , Wolfram Burgard

Reconstructing 3D human shape and pose from monocular images is challenging despite the promising results achieved by the most recent learning-based methods. The commonly occurred misalignment comes from the facts that the mapping from…

Computer Vision and Pattern Recognition · Computer Science 2020-12-08 Hongwen Zhang , Jie Cao , Guo Lu , Wanli Ouyang , Zhenan Sun

Furniture assembly is a crucial yet challenging task for robots, requiring precise dual-arm coordination where one arm manipulates parts while the other provides collaborative support and stabilization. To accomplish this task more…

Robotics · Computer Science 2026-01-19 Jiaqi Liang , Yue Chen , Qize Yu , Yan Shen , Haipeng Zhang , Hao Dong , Ruihai Wu

This paper introduces a vision-based framework for capturing and understanding human behavior in industrial assembly lines, focusing on car door manufacturing. The framework leverages advanced computer vision techniques to estimate workers'…

Computer Vision and Pattern Recognition · Computer Science 2024-09-27 Konstantinos Papoutsakis , Nikolaos Bakalos , Konstantinos Fragkoulis , Athena Zacharia , Georgia Kapetadimitri , Maria Pateraki

This paper proposes a modular framework to generate robust biped locomotion using a tight coupling between an analytical walking approach and deep reinforcement learning. This framework is composed of six main modules which are…

Robotics · Computer Science 2021-12-23 Mohammadreza Kasaei , Miguel Abreu , Nuno Lau , Artur Pereira , Luis Paulo Reis

Human motion prediction is a fundamental part of many human-robot applications. Despite the recent progress in human motion prediction, most studies simplify the problem by predicting the human motion relative to a fixed joint and/or only…

Computer Vision and Pattern Recognition · Computer Science 2022-10-04 Payam Nikdel , Mohammad Mahdavian , Mo Chen

Most state-of-the-art works in trajectory forecasting for automotive target predicting the pose and orientation of the agents in the scene. This represents a particularly useful problem, for instance in autonomous driving, but it does not…

Robotics · Computer Science 2024-10-28 Luca Paparusso , Stefano Melzi , Francesco Braghin

We present DYNARTmo, a dynamic articulatory model designed to visualize speech articulation processes in a two-dimensional midsagittal plane. The model builds upon the UK-DYNAMO framework and integrates principles of articulatory…

Computation and Language · Computer Science 2025-11-07 Bernd J. Kröger

We present a novel approach to detect, segment, and reconstruct complete textured 3D models of vehicles from a single image for autonomous driving. Our approach combines the strengths of deep learning and the elegance of traditional…

Computer Vision and Pattern Recognition · Computer Science 2020-07-17 Feixiang Lu , Zongdai Liu , Xibin Song , Dingfu Zhou , Wei Li , Hui Miao , Miao Liao , Liangjun Zhang , Bin Zhou , Ruigang Yang , Dinesh Manocha

Image-guided object assembly represents a burgeoning research topic in computer vision. This paper introduces a novel task: translating multi-view images of a structural 3D model (for example, one constructed with building blocks drawn from…

Computer Vision and Pattern Recognition · Computer Science 2024-04-26 Hongyu Yan , Yadong Mu