English
Related papers

Related papers: MV-TON: Memory-based Video Virtual Try-on network

200 papers

Video generation is important, especially in medicine, as much data is given in this form. However, video generation of high-resolution data is a very demanding task for generative models, due to the large need for memory. In this paper, we…

Image and Video Processing · Electrical Eng. & Systems 2023-11-08 Łukasz Struski , Tomasz Urbańczyk , Krzysztof Bucki , Bartłomiej Cupiał , Aneta Kaczyńska , Przemysław Spurek , Jacek Tabor

We present an end-to-end virtual try-on pipeline, that can fit different clothes on a personalized 3-D human model, reconstructed using a single RGB image. Our main idea is to construct an animatable 3-D human model and try-on different…

Computer Vision and Pattern Recognition · Computer Science 2022-10-18 Gayal Kuruppu , Bumuthu Dilshan , Shehan Samarasinghe , Nipuna Madhushan , Ranga Rodrigo

The garment transfer problem comprises two tasks: learning to separate a person's body (pose, shape, color) from their clothing (garment type, shape, style) and then generating new images of the wearer dressed in arbitrary garments. We…

Computer Vision and Pattern Recognition · Computer Science 2020-03-05 Amir Hossein Raffiee , Michael Sollami

We introduce Spatial-Temporal Memory Networks for video object detection. At its core, a novel Spatial-Temporal Memory module (STMM) serves as the recurrent computation unit to model long-term temporal appearance and motion dynamics. The…

Computer Vision and Pattern Recognition · Computer Science 2018-07-30 Fanyi Xiao , Yong Jae Lee

A key challenge in manipulation is learning a policy that can robustly generalize to diverse visual environments. A promising mechanism for learning robust policies is to leverage video generative models, which are pretrained on large-scale…

Developing Video-Grounded Dialogue Systems (VGDS), where a dialogue is conducted based on visual and audio aspects of a given video, is significantly more challenging than traditional image or text-grounded dialogue systems because (1)…

Computation and Language · Computer Science 2020-02-26 Hung Le , Doyen Sahoo , Nancy F. Chen , Steven C. H. Hoi

Virtual try-on focuses on adjusting the given clothes to fit a specific person seamlessly while avoiding any distortion of the patterns and textures of the garment. However, the clothing identity uncontrollability and training inefficiency…

Computer Vision and Pattern Recognition · Computer Science 2024-04-02 Jiazheng Xing , Chao Xu , Yijie Qian , Yang Liu , Guang Dai , Baigui Sun , Yong Liu , Jingdong Wang

A great challenge in video-language (VidL) modeling lies in the disconnection between fixed video representations extracted from image/video understanding models and downstream VidL data. Recent studies try to mitigate this disconnection…

Computer Vision and Pattern Recognition · Computer Science 2022-04-19 Tsu-Jui Fu , Linjie Li , Zhe Gan , Kevin Lin , William Yang Wang , Lijuan Wang , Zicheng Liu

Virtual try-on attracts increasing research attention as a promising way for enhancing the user experience for online cloth shopping. Though existing methods can generate impressive results, users need to provide a well-designed reference…

Computer Vision and Pattern Recognition · Computer Science 2023-05-09 Anran Lin , Nanxuan Zhao , Shuliang Ning , Yuda Qiu , Baoyuan Wang , Xiaoguang Han

This research study proposes using Generative Adversarial Networks (GAN) that incorporate a two-dimensional measure of human memorability to generate memorable or non-memorable images of scenes. The memorability of the generated images is…

Computer Vision and Pattern Recognition · Computer Science 2020-05-07 Cameron Kyle-Davidson , Adrian G. Bors , Karla K. Evans

Recent advances in Virtual Try-On (VTON) and Virtual Try-Off (VTOFF) have greatly improved photo-realistic fashion synthesis and garment reconstruction. However, existing datasets remain static, lacking instruction-driven editing for…

Computer Vision and Pattern Recognition · Computer Science 2026-04-15 Fulvio Sanguigni , Davide Lobba , Bin Ren , Marcella Cornia , Nicu Sebe , Rita Cucchiara

Recent advances in image generation and editing have opened new opportunities for virtual try-on. However, existing methods still struggle to meet complex real-world demands. We present Tstars-Tryon 1.0, a commercial-scale virtual try-on…

The growing digital landscape of fashion e-commerce calls for interactive and user-friendly interfaces for virtually trying on clothes. Traditional try-on methods grapple with challenges in adapting to diverse backgrounds, poses, and…

Human-Computer Interaction · Computer Science 2024-02-06 Justin Blalock , David Munechika , Harsha Karanth , Alec Helbling , Pratham Mehta , Seongmin Lee , Duen Horng Chau

Per-garment virtual try-on methods collect garment-specific datasets and train networks tailored to each garment to achieve superior results. However, these approaches often struggle with loose-fitting garments due to two key limitations:…

Graphics · Computer Science 2025-09-05 Zaiqiang Wu , I-Chao Shen , Takeo Igarashi

While today's video recognition systems parse snapshots or short clips accurately, they cannot connect the dots and reason across a longer range of time yet. Most existing video architectures can only process <5 seconds of a video without…

Computer Vision and Pattern Recognition · Computer Science 2022-12-02 Chao-Yuan Wu , Yanghao Li , Karttikeya Mangalam , Haoqi Fan , Bo Xiong , Jitendra Malik , Christoph Feichtenhofer

Masked video modeling~(MVM) has emerged as a highly effective pre-training strategy for visual foundation models, whereby the model reconstructs masked spatiotemporal tokens using information from visible tokens. However, a key challenge in…

Computer Vision and Pattern Recognition · Computer Science 2025-08-15 Ayush K. Rai , Kyle Min , Tarun Krishna , Feiyan Hu , Alan F. Smeaton , Noel E. O'Connor

Image-based virtual try-on aims to fit an in-shop garment into a clothed person image. To achieve this, a key step is garment warping which spatially aligns the target garment with the corresponding body parts in the person image. Prior…

Computer Vision and Pattern Recognition · Computer Science 2022-04-05 Sen He , Yi-Zhe Song , Tao Xiang

Vision-and-Language Navigation (VLN) is a task that an agent is required to follow a language instruction to navigate to the goal position, which relies on the ongoing interactions with the environment during moving. Recent…

Computer Vision and Pattern Recognition · Computer Science 2022-07-19 Chuang Lin , Yi Jiang , Jianfei Cai , Lizhen Qu , Gholamreza Haffari , Zehuan Yuan

Most virtual try-on research is motivated to serve the fashion business by generating images to demonstrate garments on studio models at a lower cost. However, virtual try-on should be a broader application that also allows customers to…

Computer Vision and Pattern Recognition · Computer Science 2024-07-18 Aiyu Cui , Jay Mahajan , Viraj Shah , Preeti Gomathinayagam , Chang Liu , Svetlana Lazebnik

Given an isolated garment image in a canonical product view and a separate image of a person, the virtual try-on task aims to generate a new image of the person wearing the target garment. Prior virtual try-on works face two major…

Computer Vision and Pattern Recognition · Computer Science 2025-05-08 Nannan Li , Kevin J. Shih , Bryan A. Plummer
‹ Prev 1 8 9 10 Next ›