English
Related papers

Related papers: Visual Persona: Foundation Model for Full-Body Hum…

200 papers

Given a person and a garment image, virtual try-on (VTO) aims to synthesize a realistic image of the person wearing the garment, while preserving their original pose and identity. Although recent VTO methods excel at visualizing garment…

Computer Vision and Pattern Recognition · Computer Science 2026-04-10 Johanna Karras , Yuanhao Wang , Yingwei Li , Ira Kemelmacher-Shlizerman

Identity-consistent generation has become an important focus in text-to-image research, with recent models achieving notable success in producing images aligned with a reference identity. Yet, the scarcity of large-scale paired datasets…

Computer Vision and Pattern Recognition · Computer Science 2025-10-17 Hengyuan Xu , Wei Cheng , Peng Xing , Yixiao Fang , Shuhan Wu , Rui Wang , Xianfang Zeng , Daxin Jiang , Gang Yu , Xingjun Ma , Yu-Gang Jiang

In this paper we predict a full 3D avatar of a person from a single image. We infer texture and geometry in the UV-space of the SMPL model using an image-to-image translation method. Given partial texture and segmentation layout maps…

Computer Vision and Pattern Recognition · Computer Science 2019-08-21 Verica Lazova , Eldar Insafutdinov , Gerard Pons-Moll

Current human image customization methods leverage Stable Diffusion (SD) for its rich semantic prior. However, since SD is not specifically designed for human-oriented generation, these methods often require extensive fine-tuning on…

Computer Vision and Pattern Recognition · Computer Science 2024-11-19 Yibin Wang , Weizhong Zhang , Cheng Jin

Text-to-image person retrieval aims to identify the target person based on a given textual description query. The primary challenge is to learn the mapping of visual and textual modalities into a common latent space. Prior works have…

Computer Vision and Pattern Recognition · Computer Science 2023-03-23 Ding Jiang , Mang Ye

Common and important applications of person identification occur at distances and viewpoints in which the face is not visible or is not sufficiently resolved to be useful. We examine body shape as a biometric across distance and viewpoint…

Computer Vision and Pattern Recognition · Computer Science 2023-05-31 Blake A. Myers , Lucas Jaggernauth , Thomas M. Metz , Matthew Q. Hill , Veda Nandan Gandi , Carlos D. Castillo , Alice J. O'Toole

Recent advances in 3D human shape estimation build upon parametric representations that model very well the shape of the naked body, but are not appropriate to represent the clothing geometry. In this paper, we present an approach to model…

Computer Vision and Pattern Recognition · Computer Science 2019-04-10 Albert Pumarola , Jordi Sanchez , Gary P. T. Choi , Alberto Sanfeliu , Francesc Moreno-Noguer

Diffusion models (DMs) have become the new trend of generative models and have demonstrated a powerful ability of conditional synthesis. Among those, text-to-image diffusion models pre-trained on large-scale image-text pairs are highly…

Computer Vision and Pattern Recognition · Computer Science 2023-03-06 Wenliang Zhao , Yongming Rao , Zuyan Liu , Benlin Liu , Jie Zhou , Jiwen Lu

Human shape editing enables controllable transformation of a person's body shape, such as thin, muscular, or overweight, while preserving pose, identity, clothing, and background. Unlike human pose editing, which has advanced rapidly, shape…

Computer Vision and Pattern Recognition · Computer Science 2025-09-26 Siddharth Khandelwal , Sridhar Kamath , Arjun Jain

The task of realistically inserting a human from a reference image into a background scene is highly challenging, requiring the model to (1) determine the correct location and poses of the person and (2) perform high-quality personalization…

Computer Vision and Pattern Recognition · Computer Science 2025-10-08 Jialu Gao , K J Joseph , Fernando De La Torre

We present BootComp, a novel framework based on text-to-image diffusion models for controllable human image generation with multiple reference garments. Here, the main bottleneck is data acquisition for training: collecting a large-scale…

Computer Vision and Pattern Recognition · Computer Science 2025-04-02 Yisol Choi , Sangkyung Kwak , Sihyun Yu , Hyungwon Choi , Jinwoo Shin

We explore the task of recognizing peoples' identities in photo albums in an unconstrained setting. To facilitate this, we introduce the new People In Photo Albums (PIPA) dataset, consisting of over 60000 instances of 2000 individuals…

Computer Vision and Pattern Recognition · Computer Science 2015-02-02 Ning Zhang , Manohar Paluri , Yaniv Taigman , Rob Fergus , Lubomir Bourdev

Many vision applications require identity consistency beyond strict biometric recognition, especially under non-frontal views or when facial cues are missing. However, conventional face recognition models enforce intra-identity invariance,…

Computer Vision and Pattern Recognition · Computer Science 2026-05-11 Yingfeng Wang , Yuxuan Xiao , Shengcai Liao

Text-to-image diffusion models have made significant advancements in generating high-quality, diverse images from text prompts. However, the inherent limitations of textual signals often prevent these models from fully capturing specific…

Computer Vision and Pattern Recognition · Computer Science 2025-03-19 Ziqiang Li , Jun Li , Lizhi Xiong , Zhangjie Fu , Zechao Li

We present a novel method for reconstructing personalized 3D human avatars with realistic animation from only a few images. Due to the large variations in body shapes, poses, and cloth types, existing methods mostly require hours of…

Computer Vision and Pattern Recognition · Computer Science 2025-04-07 Rong Wang , Fabian Prada , Ziyan Wang , Zhongshi Jiang , Chengxiang Yin , Junxuan Li , Shunsuke Saito , Igor Santesteban , Javier Romero , Rohan Joshi , Hongdong Li , Jason Saragih , Yaser Sheikh

We present Vid2Avatar-Pro, a method to create photorealistic and animatable 3D human avatars from monocular in-the-wild videos. Building a high-quality avatar that supports animation with diverse poses from a monocular video is challenging…

Computer Vision and Pattern Recognition · Computer Science 2025-03-04 Chen Guo , Junxuan Li , Yash Kant , Yaser Sheikh , Shunsuke Saito , Chen Cao

Human insertion aims to naturally place specific individuals into a target background. Although existing image editing models may have such ability, they often produce failure cases, including inappropriate human pose in new background,…

Computer Vision and Pattern Recognition · Computer Science 2026-05-11 Jie Li , Shulian Zhang , Yangyang Gao , Wenbo Li , Yulun Zhang , Yong Guo , Jian Chen

Human cognition significantly influences expressed behavior and is intrinsically tied to authentic personality traits. Personality assessment plays a pivotal role in various fields, including psychology, education, social media, etc.…

Human-Computer Interaction · Computer Science 2024-07-30 Xintong Zhang , Di Lu , Huiqi Hu , Nan Jiang , Xianhao Yu , Jinan Xu , Yujia Peng , Qing Li , Wenjuan Han

Image-based 3D Virtual Try-ON (VTON) aims to sculpt the 3D human according to person and clothes images, which is data-efficient (i.e., getting rid of expensive 3D data) but challenging. Recent text-to-3D methods achieve remarkable…

Computer Vision and Pattern Recognition · Computer Science 2024-07-24 Zhenyu Xie , Haoye Dong , Yufei Gao , Zehua Ma , Xiaodan Liang

Generation of high-quality person images is challenging, due to the sophisticated entanglements among image factors, e.g., appearance, pose, foreground, background, local details, global structures, etc. In this paper, we present a novel…

Computer Vision and Pattern Recognition · Computer Science 2020-07-20 Siyu Huang , Haoyi Xiong , Zhi-Qi Cheng , Qingzhong Wang , Xingran Zhou , Bihan Wen , Jun Huan , Dejing Dou
‹ Prev 1 3 4 5 6 7 10 Next ›