中文
相关论文

相关论文: PAD-Hand: Physics-Aware Diffusion for Hand Motion …

200 篇论文

Physics-informed deep learning has been developed as a novel paradigm for learning physical dynamics recently. While general physics-informed deep learning methods have shown early promise in learning fluid dynamics, they are difficult to…

流体动力学 · 物理学 2024-06-07 Jing Qiu , Jiancheng Huang , Xiangdong Zhang , Zeng Lin , Minglei Pan , Zengding Liu , Fen Miao

Current video diffusion models generate visually compelling content but often violate basic laws of physics, producing subtle artifacts like rubber-sheet deformations and inconsistent object motion. We introduce a frequency-domain physics…

计算机视觉与模式识别 · 计算机科学 2025-09-29 Bowen Xue , Giuseppe Claudio Guarnera , Shuang Zhao , Zahra Montazeri

Automated 3D scene generation is pivotal for applications spanning virtual reality, digital content creation, and Embodied AI. While computer graphics prioritizes aesthetic layouts, vision and robotics demand scenes that mirror real-world…

图形学 · 计算机科学 2026-03-31 Minzhang Li , Kuixiang Shao , Xuebing Li , Yuyang Jiao , Yinuo Bai , Hengan Zhou , Sixian Shen , Jiayuan Gu , Jingyi Yu

We propose a physics-based method for synthesizing dexterous hand-object interactions in a full-body setting. While recent advancements have addressed specific facets of human-object interactions, a comprehensive physics-based approach…

机器人学 · 计算机科学 2023-09-15 Jona Braun , Sammy Christen , Muhammed Kocabas , Emre Aksan , Otmar Hilliges

We propose a novel framework to reconstruct accurate appearance and geometry with neural radiance fields (NeRF) for interacting hands, enabling the rendering of photo-realistic images and videos for gesture animation from arbitrary views.…

计算机视觉与模式识别 · 计算机科学 2023-03-27 Zhiyang Guo , Wengang Zhou , Min Wang , Li Li , Houqiang Li

Predicting and generating human hand grasp over objects is critical for animation and robotic tasks. In this work, we focus on generating both the hand and objects in a grasp by a single diffusion model. Our proposed Joint Hand-Object…

计算机视觉与模式识别 · 计算机科学 2026-01-28 Jinkun Cao , Jingyuan Liu , Kris Kitani , Yi Zhou

Understanding how humans would behave during hand-object interaction is vital for applications in service robot manipulation and extended reality. To achieve this, some recent works have been proposed to simultaneously forecast hand…

计算机视觉与模式识别 · 计算机科学 2025-11-17 Junyi Ma , Jingyi Xu , Xieyuanli Chen , Hesheng Wang

Blind image restoration remains a significant challenge in low-level vision tasks. Recently, denoising diffusion models have shown remarkable performance in image synthesis. Guided diffusion models, leveraging the potent generative priors…

计算机视觉与模式识别 · 计算机科学 2024-11-20 Jun Xiao , Zihang Lyu , Hao Xie , Cong Zhang , Yakun Ju , Changjian Shui , Kin-Man Lam

We introduce a novel approach for 3D whole-body pose estimation, addressing the challenge of scale -- and deformability -- variance across body parts brought by the challenge of extending the 17 major joints on the human body to…

计算机视觉与模式识别 · 计算机科学 2025-01-06 Nermin Samet , Cédric Rommel , David Picard , Eduardo Valle

A conditional latent-diffusion based framework for solving the electromagnetic inverse scattering problem associated with microwave imaging is introduced. This generative machine-learning model explicitly mirrors the non-uniqueness of the…

图像与视频处理 · 电气工程与系统科学 2025-10-30 Shirin Chehelgami , Joe LoVetri , Vahab Khoshdel

We present a versatile latent representation that enables physically simulated character to efficiently utilize motion priors. To build a powerful motion embedding that is shared across multiple tasks, the physics controller should employ…

图形学 · 计算机科学 2025-03-18 Jinseok Bae , Jungdam Won , Donggeun Lim , Inwoo Hwang , Young Min Kim

All hand-object interaction is controlled by forces that the two bodies exert on each other, but little work has been done in modeling these underlying forces when doing pose and contact estimation from RGB/RGB-D data. Given the pose of the…

计算机视觉与模式识别 · 计算机科学 2021-08-26 Akarsh Kumar , Aditya R. Vaidya , Alexander G. Huth

Multi-fingered hands are emerging as powerful platforms for performing fine manipulation tasks, including tool use. However, environmental perturbations or execution errors can impede task performance, motivating the use of recovery…

机器人学 · 计算机科学 2025-10-09 Abhinav Kumar , Fan Yang , Sergio Aguilera Marinovic , Soshi Iba , Rana Soltani Zarrin , Dmitry Berenson

This article presents a multi-physics methodology for the numerical simulation of physical systems that involve the non-linear interaction of multi-phase reactive fluids and elastoplastic solids, inducing high strain-rates and high…

计算物理 · 物理学 2021-06-04 Tim Wallis , Philip T. Barton , Nikolaos Nikiforakis

Reconstructing high-fidelity 3D hands from egocentric monocular videos remains a challenge due to the limitations in capturing high-resolution geometry, hand-object interactions, and complex objects on hands. Additionally, existing methods…

计算机视觉与模式识别 · 计算机科学 2026-04-13 Haoyu Zhu , Yi Zhang , Lei Yao , Lap-pui Chau , Yi Wang

The malformed hands in the AI-generated images seriously affect the authenticity of the images. To refine malformed hands, existing depth-based approaches use a hand depth estimator to guide the refinement of malformed hands. Due to the…

计算机视觉与模式识别 · 计算机科学 2025-06-18 Chen-Bin Feng , Kangdao Liu , Jian Sun , Jiping Jin , Yiguo Jiang , Chi-Man Vong

We propose an approach to estimate arm and hand dynamics from monocular video by utilizing the relationship between arm and hand. Although monocular full human motion capture technologies have made great progress in recent years, recovering…

计算机视觉与模式识别 · 计算机科学 2022-03-31 Shuying Liu , Wenbin Wu , Jiaxian Wu , Yue Lin

In recent years, there has been rapid development in 3D generation models, opening up new possibilities for applications such as simulating the dynamic movements of 3D objects and customizing their behaviors. However, current 3D generative…

计算机视觉与模式识别 · 计算机科学 2024-06-12 Fangfu Liu , Hanyang Wang , Shunyu Yao , Shengjun Zhang , Jie Zhou , Yueqi Duan

In this work, we rethink the approach to video super-resolution by introducing a method based on the Diffusion Posterior Sampling framework, combined with an unconditional video diffusion transformer operating in latent space. The video…

计算机视觉与模式识别 · 计算机科学 2025-11-05 Zhihao Zhan , Wang Pang , Xiang Zhu , Yechao Bai

Physically-inspired latent force models offer an interpretable alternative to purely data driven tools for inference in dynamical systems. They carry the structure of differential equations and the flexibility of Gaussian processes,…

机器学习 · 计算机科学 2022-01-25 Jacob D. Moss , Felix L. Opolka , Bianca Dumitrascu , Pietro Lió