中文
相关论文

相关论文: CORAL: Correspondence Alignment for Improved Virtu…

200 篇论文

Learning 3D human-object interaction relation is pivotal to embodied AI and interaction modeling. Most existing methods approach the goal by learning to predict isolated interaction elements, e.g., human contact, object affordance, and…

计算机视觉与模式识别 · 计算机科学 2024-04-02 Yuhang Yang , Wei Zhai , Hongchen Luo , Yang Cao , Zheng-Jun Zha

Virtual try-on methods based on diffusion models achieve realistic effects but often require additional encoding modules, a large number of training parameters, and complex preprocessing, which increases the burden on training and…

计算机视觉与模式识别 · 计算机科学 2025-02-18 Zheng Chong , Xiao Dong , Haoxiang Li , Shiyue Zhang , Wenqing Zhang , Xujie Zhang , Hanqing Zhao , Dongmei Jiang , Xiaodan Liang

Text-driven person image generation is an emerging and challenging task in cross-modality image generation. Controllable person image generation promotes a wide range of applications such as digital human interaction and virtual try-on.…

计算机视觉与模式识别 · 计算机科学 2022-11-14 Kaiduo Zhang , Muyi Sun , Jianxin Sun , Binghao Zhao , Kunbo Zhang , Zhenan Sun , Tieniu Tan

Image-based virtual try-on systems,which fit new garments onto human portraits,are gaining research attention.An ideal pipeline should preserve the static features of clothes(like textures and logos)while also generating dynamic…

计算机视觉与模式识别 · 计算机科学 2024-01-23 Yanlong Zang , Han Yang , Jiaxu Miao , Yi Yang

Virtual try-on is a critical image synthesis task that aims to transfer clothes from one image to another while preserving the details of both humans and clothes. While many existing methods rely on Generative Adversarial Networks (GANs) to…

计算机视觉与模式识别 · 计算机科学 2023-08-14 Junhong Gou , Siyu Sun , Jianfu Zhang , Jianlou Si , Chen Qian , Liqing Zhang

Vanilla text-to-image diffusion models struggle with generating accurate human images, commonly resulting in imperfect anatomies such as unnatural postures or disproportionate limbs.Existing methods address this issue mostly by fine-tuning…

计算机视觉与模式识别 · 计算机科学 2024-03-11 Junyan Wang , Zhenhong Sun , Zhiyu Tan , Xuanbai Chen , Weihua Chen , Hao Li , Cheng Zhang , Yang Song

Vision-language models (VLMs) like CLIP have showcased a remarkable ability to extract transferable features for downstream tasks. Nonetheless, the training process of these models is usually based on a coarse-grained contrastive loss…

Text-to-image person retrieval aims to identify the target person based on a given textual description query. The primary challenge is to learn the mapping of visual and textual modalities into a common latent space. Prior works have…

计算机视觉与模式识别 · 计算机科学 2023-03-23 Ding Jiang , Mang Ye

We present an image-based VIirtual Try-On Network (VITON) without using 3D information in any form, which seamlessly transfers a desired clothing item onto the corresponding region of a person using a coarse-to-fine strategy. Conditioned…

计算机视觉与模式识别 · 计算机科学 2018-06-14 Xintong Han , Zuxuan Wu , Zhe Wu , Ruichi Yu , Larry S. Davis

In the task of reference-based image inpainting, an additional reference image is provided to restore a damaged target image to its original state. The advancement of diffusion models, particularly Stable Diffusion, allows for simple…

计算机视觉与模式识别 · 计算机科学 2025-01-07 Kuan-Hung Liu , Cheng-Kun Yang , Min-Hung Chen , Yu-Lun Liu , Yen-Yu Lin

Virtual try-on is a promising computer vision topic with a high commercial value wherein a new garment is visually worn on a person with a photo-realistic effect. Previous studies conduct their shape and content inference at one stage,…

计算机视觉与模式识别 · 计算机科学 2023-12-06 Naiyu Fang , Lemiao Qiu , Shuyou Zhang , Zili Wang , Kerui Hu

Traditional virtual try-on methods primarily focus on the garment-to-person try-on task, which requires flat garment representations. In contrast, this paper introduces a novel approach to the person-to-person try-on task. Unlike the…

计算机视觉与模式识别 · 计算机科学 2025-07-23 Zheng Wang , Xianbing Sun , Shengyi Wu , Jiahui Zhan , Jianlou Si , Chi Zhang , Liqing Zhang , Jianfu Zhang

We propose a deep learning approach for finding dense correspondences between 3D scans of people. Our method requires only partial geometric information in the form of two depth maps or partial reconstructed surfaces, works for humans in…

计算机视觉与模式识别 · 计算机科学 2016-06-28 Lingyu Wei , Qixing Huang , Duygu Ceylan , Etienne Vouga , Hao Li

Virtual 3D try-on can provide an intuitive and realistic view for online shopping and has a huge potential commercial value. However, existing 3D virtual try-on methods mainly rely on annotated 3D human shapes and garment templates, which…

计算机视觉与模式识别 · 计算机科学 2021-08-12 Fuwei Zhao , Zhenyu Xie , Michael Kampffmeyer , Haoye Dong , Songfang Han , Tianxiang Zheng , Tao Zhang , Xiaodan Liang

Vision-language models like CLIP have shown impressive capabilities in aligning images and text, but they often struggle with lengthy and detailed text descriptions because of their training focus on short and concise captions. We present…

计算机视觉与模式识别 · 计算机科学 2025-03-26 Hyungyu Choi , Young Kyun Jang , Chanho Eom

Semi-supervised learning (SSL) has been widely used to learn from both a few labeled images and many unlabeled images to overcome the scarcity of labeled samples in medical image segmentation. Most current SSL-based segmentation methods use…

计算机视觉与模式识别 · 计算机科学 2024-10-22 Xinze Li , Runlin Huang , Zhenghao Wu , Bohan Yang , Wentao Fan , Chengzhang Zhu , Weifeng Su

Achieving successful scan matching is essential for LiDAR odometry. However, in challenging environments with adverse weather conditions or repetitive geometric patterns, LiDAR odometry performance is degraded due to incorrect scan…

机器人学 · 计算机科学 2025-11-25 Jiwoo Kim , Geunsik Bae , Changseung Kim , Jinwoo Lee , Woojae Shin , Hyondong Oh

Large language models (LLMs) exhibit persistent miscalibration, especially after instruction tuning and preference alignment. Modified training objectives can improve calibration, but retraining is expensive. Inference-time steering offers…

机器学习 · 计算机科学 2026-02-06 Miranda Muqing Miao , Young-Min Cho , Lyle Ungar

We present a new method for real-time non-rigid dense correspondence between point clouds based on structured shape construction. Our method, termed Deep Point Correspondence (DPC), requires a fraction of the training data compared to…

计算机视觉与模式识别 · 计算机科学 2021-12-15 Itai Lang , Dvir Ginzburg , Shai Avidan , Dan Raviv

Virtual try-on has emerged as a pivotal task at the intersection of computer vision and fashion, aimed at digitally simulating how clothing items fit on the human body. Despite notable progress in single-image virtual try-on (VTO), current…

计算机视觉与模式识别 · 计算机科学 2025-03-12 Siqi Li , Zhengkai Jiang , Jiawei Zhou , Zhihong Liu , Xiaowei Chi , Haoqian Wang