中文
相关论文

相关论文: OmniTry: Virtual Try-On Anything without Masks

200 篇论文

Given an input video of a person and a new garment, the objective of this paper is to synthesize a new video where the person is wearing the specified garment while maintaining spatiotemporal consistency. Although significant advances have…

计算机视觉与模式识别 · 计算机科学 2024-12-19 Hung Nguyen , Quang Qui-Vinh Nguyen , Khoi Nguyen , Rang Nguyen

Given a clothing image and a person image, an image-based virtual try-on aims to generate a customized image that appears natural and accurately reflects the characteristics of the clothing image. In this work, we aim to expand the…

计算机视觉与模式识别 · 计算机科学 2023-12-05 Jeongho Kim , Gyojung Gu , Minho Park , Sunghyun Park , Jaegul Choo

Despite their impressive generative performance, latent diffusion model-based virtual try-on (VTON) methods lack faithfulness to crucial details of the clothes, such as style, pattern, and text. To alleviate these issues caused by the…

计算机视觉与模式识别 · 计算机科学 2024-05-21 Chenhui Wang , Tao Chen , Zhihao Chen , Zhizhong Huang , Taoran Jiang , Qi Wang , Hongming Shan

To enable large-scale reuse of real-world 3D assets, where garments and characters rarely share skeletons, templates, or dense correspondences, we present a fully automated virtual try-on system that dresses complex, multi-layer garments…

计算机视觉与模式识别 · 计算机科学 2026-05-18 Cong Cao , Xianhang Cheng , Jingyuan Liu , Yujian Zheng , Zhenhui Lin , Ren Li , Meriem Chkir , Hao Li

Open-vocabulary multiple object tracking aims to generalize trackers to unseen categories during training, enabling their application across a variety of real-world scenarios. However, the existing open-vocabulary tracker is constrained by…

计算机视觉与模式识别 · 计算机科学 2025-04-01 Jinyang Li , En Yu , Sijia Chen , Wenbing Tao

This paper introduces ITA-MDT, the Image-Timestep-Adaptive Masked Diffusion Transformer Framework for Image-Based Virtual Try-On (IVTON), designed to overcome the limitations of previous approaches by leveraging the Masked Diffusion…

计算机视觉与模式识别 · 计算机科学 2025-06-03 Ji Woo Hong , Tri Ton , Trung X. Pham , Gwanhyeong Koo , Sunjae Yoon , Chang D. Yoo

This paper tackles the problem of object counting in images. Existing approaches rely on extensive training data with point annotations for each object, making data collection labor-intensive and time-consuming. To overcome this, we propose…

计算机视觉与模式识别 · 计算机科学 2023-09-01 Zenglin Shi , Ying Sun , Mengmi Zhang

The automatic detection and tracking of general objects (like persons, animals or cars), text and logos in a video is crucial for many video understanding tasks, and usually real-time processing as required. We propose OmniTrack, an…

计算机视觉与模式识别 · 计算机科学 2019-10-15 Hannes Fassold , Ridouane Ghermi

Recent advances in image generation and editing have opened new opportunities for virtual try-on. However, existing methods still struggle to meet complex real-world demands. We present Tstars-Tryon 1.0, a commercial-scale virtual try-on…

Virtual try-on can significantly improve the garment shopping experiences in both online and in-store scenarios, attracting broad interest in computer vision. However, to achieve high-fidelity try-on performance, most state-of-the-art…

计算机视觉与模式识别 · 计算机科学 2024-02-06 Yunfang Niu , Dong Yi , Lingxiang Wu , Zhiwei Liu , Pengxiang Cai , Jinqiao Wang

The Diffusion model has a strong ability to generate wild images. However, the model can just generate inaccurate images with the guidance of text, which makes it very challenging to directly apply the text-guided generative model for…

计算机视觉与模式识别 · 计算机科学 2023-12-27 Shufang Zhang , Minxue Ni , Lei Wang , Wenxin Ding , Shuai Chen , Yuhong Liu

Open-vocabulary multi-object tracking (OVMOT) represents a critical new challenge involving the detection and tracking of diverse object categories in videos, encompassing both seen categories (base classes) and unseen categories (novel…

计算机视觉与模式识别 · 计算机科学 2024-10-14 Zekun Qian , Ruize Han , Junhui Hou , Linqi Song , Wei Feng

Image-based virtual try-on strives to transfer the appearance of a clothing item onto the image of a target person. Prior work focuses mainly on upper-body clothes (e.g. t-shirts, shirts, and tops) and neglects full-body or lower-body…

计算机视觉与模式识别 · 计算机科学 2022-07-14 Davide Morelli , Matteo Fincato , Marcella Cornia , Federico Landi , Fabio Cesari , Rita Cucchiara

This study discusses the critical issues of Virtual Try-On in contemporary e-commerce and the prospective metaverse, emphasizing the challenges of preserving intricate texture details and distinctive features of the target person and the…

计算机视觉与模式识别 · 计算机科学 2024-07-18 Phuong Dam , Jihoon Jeong , Anh Tran , Daeyoung Kim

Virtual Try-ON (VTON) aims to synthesis specific person images dressed in given garments, which recently receives numerous attention in online shopping scenarios. Currently, the core challenges of the VTON task mainly lie in the…

计算机视觉与模式识别 · 计算机科学 2024-10-17 Jiabao Wei , Zhiyuan Ma

Fashion image editing aims to modify a person's appearance based on a given instruction. Existing methods require auxiliary tools like segmenters and keypoint extractors, lacking a flexible and unified framework. Moreover, these methods are…

计算机视觉与模式识别 · 计算机科学 2024-10-18 Yunfang Niu , Lingxiang Wu , Dong Yi , Jie Peng , Ning Jiang , Haiying Wu , Jinqiao Wang

Virtual try-on attracts increasing research attention as a promising way for enhancing the user experience for online cloth shopping. Though existing methods can generate impressive results, users need to provide a well-designed reference…

计算机视觉与模式识别 · 计算机科学 2023-05-09 Anran Lin , Nanxuan Zhao , Shuliang Ning , Yuda Qiu , Baoyuan Wang , Xiaoguang Han

Virtual try-on technology has become increasingly important in the fashion and retail industries, enabling the generation of high-fidelity garment images that adapt seamlessly to target human models. While existing methods have achieved…

计算机视觉与模式识别 · 计算机科学 2025-10-30 Ming Meng , Qi Dong , Jiajie Li , Zhe Zhu , Xingyu Wang , Zhaoxin Fan , Wei Zhao , Wenjun Wu

We present a conceptually simple, flexible, and general framework for object instance segmentation. Our approach efficiently detects objects in an image while simultaneously generating a high-quality segmentation mask for each instance. The…

计算机视觉与模式识别 · 计算机科学 2018-01-25 Kaiming He , Georgia Gkioxari , Piotr Dollár , Ross Girshick

Diffusion models have led to the revolutionizing of generative modeling in numerous image synthesis tasks. Nevertheless, it is not trivial to directly apply diffusion models for synthesizing an image of a target person wearing a given…

计算机视觉与模式识别 · 计算机科学 2024-09-13 Siqi Wan , Yehao Li , Jingwen Chen , Yingwei Pan , Ting Yao , Yang Cao , Tao Mei