中文
相关论文

相关论文: LieHMR: Autoregressive Human Mesh Recovery with $S…

200 篇论文

Diffusion-based generative models represent the current state-of-the-art for image generation. However, standard diffusion models are based on Euclidean geometry and do not translate directly to manifold-valued data. In this work, we…

机器学习 · 计算机科学 2023-12-20 Yesukhei Jagvaral , Francois Lanusse , Rachel Mandelbaum

Recent generative methods for single-shot high dynamic range (HDR) image reconstruction show promising results, but often struggle with preserving fidelity to the input image. They require separate models to handle highlights and shadows,…

计算机视觉与模式识别 · 计算机科学 2026-05-13 Chinmay Talegaonkar , Jinshi He , Christopher McKenna , Nicholas Antipa

Detailed and photorealistic 3D human modeling is essential for various applications and has seen tremendous progress. However, full-body reconstruction from a monocular RGB image remains challenging due to the ill-posed nature of the…

计算机视觉与模式识别 · 计算机科学 2025-03-25 Peng Li , Wangguandong Zheng , Yuan Liu , Tao Yu , Yangguang Li , Xingqun Qi , Xiaowei Chi , Siyu Xia , Yan-Pei Cao , Wei Xue , Wenhan Luo , Yike Guo

The end-to-end Human Mesh Recovery (HMR) approach has been successfully used for 3D body reconstruction. However, most HMR-based frameworks reconstruct human body by directly learning mesh parameters from images or videos, while lacking…

计算机视觉与模式识别 · 计算机科学 2021-03-19 Tianyu Luan , Yali Wang , Junhao Zhang , Zhe Wang , Zhipeng Zhou , Yu Qiao

Single image pose estimation is a fundamental problem in many vision and robotics tasks, and existing deep learning approaches suffer by not completely modeling and handling: i) uncertainty about the predictions, and ii) symmetric objects…

计算机视觉与模式识别 · 计算机科学 2022-07-05 Kieran Murphy , Carlos Esteves , Varun Jampani , Srikumar Ramalingam , Ameesh Makadia

Recent years have witnessed a trend of the deep integration of the generation and reconstruction paradigms. In this paper, we extend the ability of controllable generative models for a more comprehensive hand mesh recovery task: direct hand…

计算机视觉与模式识别 · 计算机科学 2024-06-04 Mengcheng Li , Hongwen Zhang , Yuxiang Zhang , Ruizhi Shao , Tao Yu , Yebin Liu

We present a novel diffusion-based approach for coherent 3D scene reconstruction from a single RGB image. Our method utilizes an image-conditioned 3D scene diffusion model to simultaneously denoise the 3D poses and geometries of all objects…

计算机视觉与模式识别 · 计算机科学 2024-12-16 Manuel Dahnert , Angela Dai , Norman Müller , Matthias Nießner

In recent years, point cloud perception tasks have been garnering increasing attention. This paper presents the first attempt to estimate 3D human body mesh from sparse LiDAR point clouds. We found that the major challenge in estimating…

计算机视觉与模式识别 · 计算机科学 2025-10-03 Bohao Fan , Wenzhao Zheng , Jianjiang Feng , Jie Zhou

One of the mainstream schemes for 2D human pose estimation (HPE) is learning keypoints heatmaps by a neural network. Existing methods typically improve the quality of heatmaps by customized architectures, such as high-resolution…

计算机视觉与模式识别 · 计算机科学 2023-06-30 Zhongwei Qiu , Qiansheng Yang , Jian Wang , Xiyu Wang , Chang Xu , Dongmei Fu , Kun Yao , Junyu Han , Errui Ding , Jingdong Wang

Recently, diffusion-based methods for monocular 3D human pose estimation have achieved state-of-the-art (SOTA) performance by directly regressing the 3D joint coordinates from the 2D pose sequence. Although some methods decompose the task…

计算机视觉与模式识别 · 计算机科学 2025-07-01 Qingyuan Cai , Xuecai Hu , Saihui Hou , Li Yao , Yongzhen Huang

Denoising diffusion probabilistic models that were initially proposed for realistic image generation have recently shown success in various perception tasks (e.g., object detection and image segmentation) and are increasingly gaining…

计算机视觉与模式识别 · 计算机科学 2023-08-08 Runyang Feng , Yixing Gao , Tze Ho Elden Tse , Xueqing Ma , Hyung Jin Chang

Digital imaging aims to replicate realistic scenes, but Low Dynamic Range (LDR) cameras cannot represent the wide dynamic range of real scenes, resulting in under-/overexposed images. This paper presents a deep learning-based approach for…

计算机视觉与模式识别 · 计算机科学 2023-07-07 Dwip Dalal , Gautam Vashishtha , Prajwal Singh , Shanmuganathan Raman

A long-standing goal of 3D human reconstruction is to create lifelike and fully detailed 3D humans from single-view images. The main challenge lies in inferring unknown body shapes, appearances, and clothing details in areas not visible in…

计算机视觉与模式识别 · 计算机科学 2024-04-02 Hsuan-I Ho , Jie Song , Otmar Hilliges

Monocular 3D human pose and shape estimation is an ill-posed problem since multiple 3D solutions can explain a 2D image of a subject. Recent approaches predict a probability distribution over plausible 3D pose and shape parameters…

计算机视觉与模式识别 · 计算机科学 2023-05-12 Akash Sengupta , Ignas Budvytis , Roberto Cipolla

Accurate 3D human pose estimation remains a critical yet unresolved challenge, requiring both temporal coherence across frames and fine-grained modeling of joint relationships. However, most existing methods rely solely on geometric cues…

计算机视觉与模式识别 · 计算机科学 2025-11-13 Jerrin Bright , Yuhao Chen , John S. Zelek

This paper addresses the problem of 3D human body shape and pose estimation from RGB images. Some recent approaches to this task predict probability distributions over human body model parameters conditioned on the input images. This is…

计算机视觉与模式识别 · 计算机科学 2021-12-01 Akash Sengupta , Ignas Budvytis , Roberto Cipolla

Reconstructing the 3D shape of an object from a single RGB image is a long-standing and highly challenging problem in computer vision. In this paper, we propose a novel method for single-image 3D reconstruction which generates a sparse…

计算机视觉与模式识别 · 计算机科学 2023-02-24 Luke Melas-Kyriazi , Christian Rupprecht , Andrea Vedaldi

Diffusion models are a special type of generative model, capable of synthesising new data from a learnt distribution. We introduce DISPR, a diffusion-based model for solving the inverse problem of three-dimensional (3D) cell shape…

计算机视觉与模式识别 · 计算机科学 2023-03-15 Dominik J. E. Waibel , Ernst Röell , Bastian Rieck , Raja Giryes , Carsten Marr

Text-driven person image generation is an emerging and challenging task in cross-modality image generation. Controllable person image generation promotes a wide range of applications such as digital human interaction and virtual try-on.…

计算机视觉与模式识别 · 计算机科学 2022-11-14 Kaiduo Zhang , Muyi Sun , Jianxin Sun , Binghao Zhao , Kunbo Zhang , Zhenan Sun , Tieniu Tan

Universal image restoration is a practical and potential computer vision task for real-world applications. The main challenge of this task is handling the different degradation distributions at once. Existing methods mainly utilize…

计算机视觉与模式识别 · 计算机科学 2024-03-19 Dian Zheng , Xiao-Ming Wu , Shuzhou Yang , Jian Zhang , Jian-Fang Hu , Wei-Shi Zheng