English
Related papers

Related papers: Fashion-VDM: Video Diffusion Model for Virtual Try…

200 papers

We present a data-driven method for learning to generate animations of 3D garments using a 2D image diffusion model. In contrast to existing methods, typically based on fully connected networks, graph neural networks, or generative…

Computer Vision and Pattern Recognition · Computer Science 2025-03-25 Raquel Vidaurre , Elena Garces , Dan Casas

Recent advances in the diffusion models have significantly improved text-to-image generation. However, generating videos from text is a more challenging task than generating images from text, due to the much larger dataset and higher…

Computer Vision and Pattern Recognition · Computer Science 2024-12-31 Taegyeong Lee , Soyeong Kwon , Taehwan Kim

A virtual try-on method takes a product image and an image of a model and produces an image of the model wearing the product. Most methods essentially compute warps from the product image to the model image and combine using image…

Computer Vision and Pattern Recognition · Computer Science 2020-03-30 Kedan Li , Min Jin Chong , Jingen Liu , David Forsyth

Virtual 3D try-on can provide an intuitive and realistic view for online shopping and has a huge potential commercial value. However, existing 3D virtual try-on methods mainly rely on annotated 3D human shapes and garment templates, which…

Computer Vision and Pattern Recognition · Computer Science 2021-08-12 Fuwei Zhao , Zhenyu Xie , Michael Kampffmeyer , Haoye Dong , Songfang Han , Tianxiang Zheng , Tao Zhang , Xiaodan Liang

Recent diffusion-based approaches have made significant advances in image-based virtual try-on, enabling more realistic and end-to-end garment synthesis. However, most existing methods remain constrained by their reliance on exhibition…

Computer Vision and Pattern Recognition · Computer Science 2026-04-15 Jinxi Liu , Zijian He , Guangrun Wang , Guanbin Li , Liang Lin

Text-to-video generation aims to produce a video based on a given prompt. Recently, several commercial video models have been able to generate plausible videos with minimal noise, excellent details, and high aesthetic scores. However, these…

Computer Vision and Pattern Recognition · Computer Science 2024-01-18 Haoxin Chen , Yong Zhang , Xiaodong Cun , Menghan Xia , Xintao Wang , Chao Weng , Ying Shan

Image-based virtual try-on aims to transfer target in-shop clothing to a dressed model image, the objectives of which are totally taking off original clothing while preserving the contents outside of the try-on area, naturally wearing…

Computer Vision and Pattern Recognition · Computer Science 2024-09-24 Dan Song , Xuanpu Zhang , Jianhao Zeng , Pengxin Zhan , Qingguo Chen , Weihua Luo , An-An Liu

Fashionable image generation aims to synthesize images of diverse fashion prevalent around the globe, helping fashion designers in real-time visualization by giving them a basic customized structure of how a specific design preference would…

Computer Vision and Pattern Recognition · Computer Science 2023-06-14 Krishna Sri Ipsit Mantri , Nevasini Sasikumar

To address the larger computation and storage requirements associated with large video datasets, video dataset distillation aims to capture spatial and temporal information in a significantly smaller dataset, such that training on the…

Computer Vision and Pattern Recognition · Computer Science 2025-07-31 Kunyang Li , Jeffrey A Chan Santiago , Sarinda Dhanesh Samarasinghe , Gaowen Liu , Mubarak Shah

Image-based virtual try-on aims to fit an in-shop garment into a clothed person image. To achieve this, a key step is garment warping which spatially aligns the target garment with the corresponding body parts in the person image. Prior…

Computer Vision and Pattern Recognition · Computer Science 2022-04-05 Sen He , Yi-Zhe Song , Tao Xiang

Virtual try-on system under arbitrary human poses has huge application potential, yet raises quite a lot of challenges, e.g. self-occlusions, heavy misalignment among diverse poses, and diverse clothes textures. Existing methods aim at…

Computer Vision and Pattern Recognition · Computer Science 2019-03-01 Haoye Dong , Xiaodan Liang , Bochao Wang , Hanjiang Lai , Jia Zhu , Jian Yin

Virtual Try-On (VTON) aims to synthesize photorealistic images of garments precisely aligned with a person's body and pose. Current diffusion-based methods, however, face a fundamental trade-off between structural integrity and textural…

Computer Vision and Pattern Recognition · Computer Science 2026-05-15 Yixin Liu , Baihong Qian , Jinglin Jiang , Jeffery Wu , Yan Chen , Wei Wang , Yida Wang , Lanqing Yang , Guangtao Xue

We present an image-based VIirtual Try-On Network (VITON) without using 3D information in any form, which seamlessly transfers a desired clothing item onto the corresponding region of a person using a coarse-to-fine strategy. Conditioned…

Computer Vision and Pattern Recognition · Computer Science 2018-06-14 Xintong Han , Zuxuan Wu , Zhe Wu , Ruichi Yu , Larry S. Davis

Recent advances in 4D generation mainly focus on generating 4D content by distilling pre-trained text or single-view image-conditioned models. It is inconvenient for them to take advantage of various off-the-shelf 3D assets with multi-view…

Computer Vision and Pattern Recognition · Computer Science 2024-09-10 Yanqin Jiang , Chaohui Yu , Chenjie Cao , Fan Wang , Weiming Hu , Jin Gao

Diffusion models have achieved great success in image generation. However, when leveraging this idea for video generation, we face significant challenges in maintaining the consistency and continuity across video frames. This is mainly…

Computer Vision and Pattern Recognition · Computer Science 2024-03-25 Haoran Lang , Yuxuan Ge , Zheng Tian

Video Virtual Try-on aims to seamlessly transfer a reference garment onto a target person in a video while preserving both visual fidelity and temporal coherence. Existing methods typically rely on inpainting masks to define the try-on…

Computer Vision and Pattern Recognition · Computer Science 2025-07-22 Tianyu Chang , Xiaohao Chen , Zhichao Wei , Xuanpu Zhang , Qing-Guo Chen , Weihua Luo , Peipei Song , Xun Yang

Text-to-video diffusion models have advanced video generation significantly. However, customizing these models to generate videos with tailored motions presents a substantial challenge. In specific, they encounter hurdles in (a) accurately…

Computer Vision and Pattern Recognition · Computer Science 2023-12-05 Hyeonho Jeong , Geon Yeong Park , Jong Chul Ye

We present DreamPose, a diffusion-based method for generating animated fashion videos from still images. Given an image and a sequence of human body poses, our method synthesizes a video containing both human and fabric motion. To achieve…

Computer Vision and Pattern Recognition · Computer Science 2023-11-01 Johanna Karras , Aleksander Holynski , Ting-Chun Wang , Ira Kemelmacher-Shlizerman

Most virtual try-on research is motivated to serve the fashion business by generating images to demonstrate garments on studio models at a lower cost. However, virtual try-on should be a broader application that also allows customers to…

Computer Vision and Pattern Recognition · Computer Science 2024-07-18 Aiyu Cui , Jay Mahajan , Viraj Shah , Preeti Gomathinayagam , Chang Liu , Svetlana Lazebnik

We introduce the Virtual Fitting Room (VFR), a novel video generative model that produces arbitrarily long virtual try-on videos. Our VFR models long video generation tasks as an auto-regressive, segment-by-segment generation process,…

Computer Vision and Pattern Recognition · Computer Science 2025-09-05 Jun-Kun Chen , Aayush Bansal , Minh Phuoc Vo , Yu-Xiong Wang