中文
相关论文

相关论文: CORAL: Correspondence Alignment for Improved Virtu…

200 篇论文

Virtual try-on (VITON) aims to generate realistic images of a person wearing a target garment, requiring precise garment alignment in try-on regions and faithful preservation of identity and background in non-try-on regions. While latent…

计算机视觉与模式识别 · 计算机科学 2026-04-15 Junseo Park , Hyeryung Jang

Virtual try-on seeks to generate photorealistic images of individuals in desired garments, a task that must simultaneously preserve personal identity and garment fidelity for practical use in fashion retail and personalization. However,…

计算机视觉与模式识别 · 计算机科学 2025-08-13 Ankan Deria , Dwarikanath Mahapatra , Behzad Bozorgtabar , Mohna Chakraborty , Snehashis Chakraborty , Sudipta Roy

Image-based Virtual Try-On (VITON) aims to transfer an in-shop garment image onto a target person. While existing methods focus on warping the garment to fit the body pose, they often overlook the synthesis quality around the garment-skin…

计算机视觉与模式识别 · 计算机科学 2023-12-07 xujie zhang , Xiu Li , Michael Kampffmeyer , Xin Dong , Zhenyu Xie , Feida Zhu , Haoye Dong , Xiaodan Liang

The goal of image-based virtual try-on is to generate an image of the target person naturally wearing the given clothing. However, existing methods solely focus on the frontal try-on using the frontal clothing. When the views of the…

计算机视觉与模式识别 · 计算机科学 2025-01-07 Haoyu Wang , Zhilu Zhang , Donglin Di , Shiliang Zhang , Wangmeng Zuo

Image-based 3D Virtual Try-ON (VTON) aims to sculpt the 3D human according to person and clothes images, which is data-efficient (i.e., getting rid of expensive 3D data) but challenging. Recent text-to-3D methods achieve remarkable…

计算机视觉与模式识别 · 计算机科学 2024-07-24 Zhenyu Xie , Haoye Dong , Yufei Gao , Zehua Ma , Xiaodan Liang

We present OOTDiffusion, a novel network architecture for realistic and controllable image-based virtual try-on (VTON). We leverage the power of pretrained latent diffusion models, designing an outfitting UNet to learn the garment detail…

计算机视觉与模式识别 · 计算机科学 2024-03-08 Yuhao Xu , Tao Gu , Weifeng Chen , Chengcai Chen

Conversational recommender systems (CRSs) are designed to suggest the target item that the user is likely to prefer through multi-turn conversations. Recent studies stress that capturing sentiments in user conversations improves…

信息检索 · 计算机科学 2025-07-30 Heejin Kook , Junyoung Kim , Seongmin Park , Jongwuk Lee

Retrieval-Augmented Generation (RAG) has become a powerful paradigm for enhancing large language models (LLMs) through external knowledge retrieval. Despite its widespread attention, existing academic research predominantly focuses on…

Virtual try-on (VTON) transfers a target clothing image to a reference person, where clothing fidelity is a key requirement for downstream e-commerce applications. However, existing VTON methods still fall short in high-fidelity try-on due…

计算机视觉与模式识别 · 计算机科学 2024-11-05 Han Yang , Yanlong Zang , Ziwei Liu

Replicating In-Context Learning (ICL) in computer vision remains challenging due to task heterogeneity. We propose \textbf{VIRAL}, a framework that elicits visual reasoning from a pre-trained image editing model by formulating ICL as…

计算机视觉与模式识别 · 计算机科学 2026-02-04 Zhiwen Li , Zhongjie Duan , Jinyan Ye , Cen Chen , Daoyuan Chen , Yaliang Li , Yingda Chen

The garment-to-person virtual try-on (VTON) task, which aims to generate fitting images of a person wearing a reference garment, has made significant strides. However, obtaining a standard garment is often more challenging than using the…

计算机视觉与模式识别 · 计算机科学 2025-02-04 Le Shen , Yanting Kang , Rong Huang , Zhijie Wang

Text-driven 3D editing seeks to modify 3D scenes according to textual descriptions, and most existing approaches tackle this by adapting pre-trained 2D image editors to multi-view inputs. However, without explicit control over multi-view…

计算机视觉与模式识别 · 计算机科学 2026-02-20 Zhe Zhu , Honghua Chen , Peng Li , Mingqiang Wei

Multilingual retrieval-augmented generation (mRAG) is often implemented within a fixed retrieval space, typically via query or document translation or multilingual embedding vector representations. However, this approach may be inadequate…

计算与语言 · 计算机科学 2026-04-29 Nayeon Lee , Jiwoo Song , Byeongcheol Kang

In this paper, we introduce D$^4$-VTON, an innovative solution for image-based virtual try-on. We address challenges from previous studies, such as semantic inconsistencies before and after garment warping, and reliance on static,…

计算机视觉与模式识别 · 计算机科学 2024-07-23 Zhaotong Yang , Zicheng Jiang , Xinzhe Li , Huiyu Zhou , Junyu Dong , Huaidong Zhang , Yong Du

Virtual Try-On (VTON) has seen rapid advancements, providing a strong foundation for generative fashion tasks. However, the inverse problem, Virtual Try-Off (VTOFF)-aimed at reconstructing the canonical garment from a draped-on…

计算机视觉与模式识别 · 计算机科学 2026-04-13 Loc-Phat Truong , Meysam Madadi , Sergio Escalera

This work aims to address a novel Customized Virtual Try-ON (Cu-VTON) task, enabling the superimposition of a specified garment onto a model that can be customized in terms of appearance, posture, and additional attributes. Compared with…

计算机视觉与模式识别 · 计算机科学 2026-02-02 Zhijing Yang , Weiwei Zhang , Mingliang Yang , Siyuan Peng , Yukai Shi , Junpeng Tan , Tianshui Chen , Liruo Zhong

Image-based virtual try-on systems for fitting new in-shop clothes into a person image have attracted increasing research attention, yet is still challenging. A desirable pipeline should not only transform the target clothes into the most…

计算机视觉与模式识别 · 计算机科学 2018-09-13 Bochao Wang , Huabin Zheng , Xiaodan Liang , Yimin Chen , Liang Lin , Meng Yang

Virtual try-on, a rapidly evolving field in computer vision, is transforming e-commerce by improving customer experiences through precise garment warping and seamless integration onto the human body. While existing methods such as TPS and…

计算机视觉与模式识别 · 计算机科学 2024-06-05 Sanhita Pathak , Vinay Kaushik , Brejesh Lall

Virtual Try-On (VTON) is the task of synthesizing an image of a person wearing a target garment, conditioned on a person image and a garment image. While diffusion-based VTON models featuring a Dual UNet architecture demonstrate superior…

计算机视觉与模式识别 · 计算机科学 2025-11-25 Kihyun Na , Jinyoung Choi , Injung Kim

The long-tail recommendation is a challenging task for traditional recommender systems, due to data sparsity and data imbalance issues. The recent development of large language models (LLMs) has shown their abilities in complex reasoning,…

信息检索 · 计算机科学 2024-03-12 Junda Wu , Cheng-Chun Chang , Tong Yu , Zhankui He , Jianing Wang , Yupeng Hou , Julian McAuley