中文
相关论文

相关论文: Human Body Restoration with One-Step Diffusion Mod…

200 篇论文

This paper presents a novel framework to recover \emph{detailed} avatar from a single image. It is a challenging task due to factors such as variations in human shapes, body poses, texture, and viewpoints. Prior methods typically attempt to…

计算机视觉与模式识别 · 计算机科学 2021-08-21 Hao Zhu , Xinxin Zuo , Haotian Yang , Sen Wang , Xun Cao , Ruigang Yang

Rectified Flow text-to-image models surpass diffusion models in image quality and text alignment, but adapting ReFlow for real-image editing remains challenging. We propose a new real-image editing method for ReFlow by analyzing the…

计算机视觉与模式识别 · 计算机科学 2025-07-03 Jimyeong Kim , Jungwon Park , Yeji Song , Nojun Kwak , Wonjong Rhee

Prior human parsing models are limited to parsing humans into classes pre-defined in the training data, which is not flexible to generalize to unseen classes, e.g., new clothing in fashion analysis. In this paper, we propose a new problem…

计算机视觉与模式识别 · 计算机科学 2021-05-10 Haoyu He , Jing Zhang , Bhavani Thuraisingham , Dacheng Tao

In this work, we enhance a professional end-to-end volumetric video production pipeline to achieve high-fidelity human body reconstruction using only passive cameras. While current volumetric video approaches estimate depth maps using…

计算机视觉与模式识别 · 计算机科学 2022-03-01 Decai Chen , Markus Worchel , Ingo Feldmann , Oliver Schreer , Peter Eisert

We introduce HART, a unified framework for sparse-view human reconstruction. Given a small set of uncalibrated RGB images of a person as input, it outputs a watertight clothed mesh, the aligned SMPL-X body mesh, and a Gaussian-splat…

计算机视觉与模式识别 · 计算机科学 2025-10-01 Xiyi Chen , Shaofei Wang , Marko Mihajlovic , Taewon Kang , Sergey Prokudin , Ming Lin

Diffusion models have achieved impressive success in high-fidelity image generation but suffer from slow sampling due to their inherently iterative denoising process. While recent one-step methods accelerate inference by learning direct…

机器学习 · 计算机科学 2025-10-15 Hanru Bai , Weiyang Ding , Difan Zou

Recent advances in diffusion models have led to impressive image generation capabilities, but aligning these models with human preferences remains challenging. Reward-based fine-tuning using models trained on human feedback improves…

计算机视觉与模式识别 · 计算机科学 2025-05-29 Dmitrii Sorokin , Maksim Nakhodnov , Andrey Kuznetsov , Aibek Alanov

Recent diffusion-based one-step methods have shown remarkable progress in the field of image super-resolution, yet they remain constrained by three critical limitations: (1) inferior fidelity performance caused by the information loss from…

计算机视觉与模式识别 · 计算机科学 2025-12-17 Hao Chen , Junyang Chen , Jinshan Pan , Jiangxin Dong

We present Multi-HMR, a strong sigle-shot model for multi-person 3D human mesh recovery from a single RGB image. Predictions encompass the whole body, i.e., including hands and facial expressions, using the SMPL-X parametric model and 3D…

计算机视觉与模式识别 · 计算机科学 2024-07-25 Fabien Baradel , Matthieu Armando , Salma Galaaoui , Romain Brégier , Philippe Weinzaepfel , Grégory Rogez , Thomas Lucas

Garment restoration, the inverse of virtual try-on task, focuses on restoring standard garment from a person image, requiring accurate capture of garment details. However, existing methods often fail to preserve the identity of the garment…

计算机视觉与模式识别 · 计算机科学 2024-12-17 Le Shen , Rong Huang , Zhijie Wang

Despite the significant progress in diffusion prior-based image restoration, most existing methods apply uniform processing to the entire image, lacking the capability to perform region-customized image restoration according to user…

计算机视觉与模式识别 · 计算机科学 2025-04-01 Shuaizheng Liu , Jianqi Ma , Lingchen Sun , Xiangtao Kong , Lei Zhang

Nuclei segmentation is a fundamental but challenging task in the quantitative analysis of histopathology images. Although fully-supervised deep learning-based methods have made significant progress, a large number of labeled images are…

图像与视频处理 · 电气工程与系统科学 2024-01-22 Xinyi Yu , Guanbin Li , Wei Lou , Siqi Liu , Xiang Wan , Yan Chen , Haofeng Li

We propose a one-step person detector for topview omnidirectional indoor scenes based on convolutional neural networks (CNNs). While state of the art person detectors reach competitive results on perspective images, missing CNN…

计算机视觉与模式识别 · 计算机科学 2022-04-15 Jingrui Yu , Roman Seidel , Gangolf Hirtz

Regression-based methods have shown high efficiency and effectiveness for multi-view human mesh recovery. The key components of a typical regressor lie in the feature extraction of input views and the fusion of multi-view features. In this…

计算机视觉与模式识别 · 计算机科学 2023-12-19 Kai Jia , Hongwen Zhang , Liang An , Yebin Liu

Diffusion models equipped with language models demonstrate excellent controllability in image generation tasks, allowing image processing to adhere to human instructions. However, the lack of diverse instruction-following data hampers the…

计算机视觉与模式识别 · 计算机科学 2024-10-11 Yongsheng Yu , Ziyun Zeng , Hang Hua , Jianlong Fu , Jiebo Luo

Restoring images afflicted by complex real-world degradations remains challenging, as conventional methods often fail to adapt to the unique mixture and severity of artifacts present. This stems from a reliance on indirect cues which poorly…

计算机视觉与模式识别 · 计算机科学 2025-04-18 Xin Su , Chen Wu , Yu Zhang , Chen Lyu , Zhuoran Zheng

Recent advances in few-step diffusion models have demonstrated their efficiency and effectiveness by shortcutting the probabilistic paths of diffusion models, especially in training one-step diffusion models from scratch (\emph{a.k.a.}…

机器学习 · 计算机科学 2026-02-03 Haitao Lin , Peiyan Hu , Minsi Ren , Zhifeng Gao , Zhi-Ming Ma , Guolin ke , Tailin Wu , Stan Z. Li

To date, little attention has been given to multi-view 3D human mesh estimation, despite real-life applicability (e.g., motion capture, sport analysis) and robustness to single-view ambiguities. Existing solutions typically suffer from poor…

计算机视觉与模式识别 · 计算机科学 2022-12-13 Xuan Gong , Liangchen Song , Meng Zheng , Benjamin Planche , Terrence Chen , Junsong Yuan , David Doermann , Ziyan Wu

Generating higher-resolution human-centric scenes with details and controls remains a challenge for existing text-to-image diffusion models. This challenge stems from limited training image size, text encoder capacity (limited tokens), and…

计算机视觉与模式识别 · 计算机科学 2024-04-09 Gwanghyun Kim , Hayeon Kim , Hoigi Seo , Dong Un Kang , Se Young Chun

We consider the problem of obese human mesh recovery, i.e., fitting a parametric human mesh to images of obese people. Despite obese person mesh fitting being an important problem with numerous applications (e.g., healthcare), much recent…

计算机视觉与模式识别 · 计算机科学 2021-07-14 Ren Li , Meng Zheng , Srikrishna Karanam , Terrence Chen , Ziyan Wu