中文
相关论文

相关论文: MMFashion: An Open-Source Toolbox for Visual Fashi…

200 篇论文

Fashion intelligence spans multiple tasks, i.e., retrieval, recommendation, recognition, and dialogue, yet remains hindered by fragmented supervision and incomplete fashion annotations. These limitations jointly restrict the formation of…

计算机视觉与模式识别 · 计算机科学 2026-03-04 Zhengwei Yang , Andi Long , Hao Li , Zechao Hu , Kui Jiang , Zheng Wang

The fashion domain encompasses a variety of real-world multimodal tasks, including multimodal retrieval and multimodal generation. The rapid advancements in artificial intelligence generated content, particularly in technologies like large…

计算机视觉与模式识别 · 计算机科学 2024-10-15 Xiangyu Zhao , Yuehan Zhang , Wenlong Zhang , Xiao-Ming Wu

Existing 4D human datasets fall short for fashion-specific research, lacking either realistic garment dynamics or task-specific annotations. Synthetic datasets suffer from a realism gap, whereas real-world captures lack the detailed…

计算机视觉与模式识别 · 计算机科学 2026-03-10 Hunor Laczkó , Libang Jia , Loc-Phat Truong , Diego Hernández , Sergio Escalera , Jordi Gonzalez , Meysam Madadi

Fashion-focused artificial intelligence has rapidly advanced in recent years, driven by deep learning and its deployment in recommender systems, detection, retrieval, and analytics. Yet several consumer-facing domains remain comparatively…

计算机视觉与模式识别 · 计算机科学 2026-03-20 Laila Khalid , Wei Gong

Fashion styling and personalized recommendations are pivotal in modern retail, contributing substantial economic value in the fashion industry. With the advent of vision-language models (VLM), new opportunities have emerged to enhance…

计算机视觉与模式识别 · 计算机科学 2025-04-28 Kaicheng Pang , Xingxing Zou , Waikeung Wong

Personalized generative recommender systems have emerged as a promising solution for fashion recommendation. However, existing methods primarily rely on implicit visual embeddings from historical interactions, which often contain…

信息检索 · 计算机科学 2026-05-19 Mingzhe Yu , Lei Wu , Qianru Sun , Yunshan Ma

The fashion industry has diverse applications in multi-modal image generation and editing. It aims to create a desired high-fidelity image with the multi-modal conditional signal as guidance. Most existing methods learn different condition…

计算机视觉与模式识别 · 计算机科学 2022-05-25 Zhikang Li , Huiling Zhou , Shuai Bai , Peike Li , Chang Zhou , Hongxia Yang

With advances in foundational and vision-language models, and effective fine-tuning techniques, a large number of both general and special-purpose models have been developed for a variety of visual tasks. Despite the flexibility and…

计算机视觉与模式识别 · 计算机科学 2024-12-25 Wan-Cyuan Fan , Tanzila Rahman , Leonid Sigal

DORAEMON is an open-source PyTorch library that unifies visual object modeling and representation learning across diverse scales. A single YAML-driven workflow covers classification, retrieval and metric learning; more than 1000 pretrained…

计算机视觉与模式识别 · 计算机科学 2025-11-07 Ke Du , Yimin Peng , Chao Gao , Fan Zhou , Siqiao Xue

In this paper, we propose a multimodal search engine that combines visual and textual cues to retrieve items from a multimedia database aesthetically similar to the query. The goal of our engine is to enable intuitive retrieval of fashion…

计算机视觉与模式识别 · 计算机科学 2019-02-21 Ivona Tautkute , Tomasz Trzcinski , Aleksander Skorupa , Lukasz Brocki , Krzysztof Marasek

Composing fashion outfits involves deep understanding of fashion standards while incorporating creativity for choosing multiple fashion items (e.g., Jewelry, Bag, Pants, Dress). In fashion websites, popular or high-quality fashion outfits…

多媒体 · 计算机科学 2017-04-18 Yuncheng Li , LiangLiang Cao , Jiang Zhu , Jiebo Luo

In this paper, we present LookBench (We use the term "look" to reflect retrieval that mirrors how people shop -- finding the exact item, a close substitute, or a visually consistent alternative.), a live, holistic and challenging benchmark…

计算机视觉与模式识别 · 计算机科学 2026-04-14 Gensmo. ai , Chao Gao , Siqiao Xue , Jiwen Fu , Tingyi Gu , Shanshan Li , Fan Zhou

This paper introduces MMTryon, a multi-modal multi-reference VIrtual Try-ON (VITON) framework, which can generate high-quality compositional try-on results by taking a text instruction and multiple garment images as inputs. Our MMTryon…

计算机视觉与模式识别 · 计算机科学 2024-11-21 Xujie Zhang , Ente Lin , Xiu Li , Yuxuan Luo , Michael Kampffmeyer , Xin Dong , Xiaodan Liang

Modern time series analysis demands frameworks that are flexible, efficient, and extensible. However, many existing Python libraries exhibit limitations in modularity and in their native support for irregular, multi-source, or sparse data.…

机器学习 · 计算机科学 2025-08-27 Zhijin Wang , Senzhen Wu , Yue Hu , Xiufeng Liu

We present an open-source toolbox, named MMRotate, which provides a coherent algorithm framework of training, inferring, and evaluation for the popular rotated object detection algorithm based on deep learning. MMRotate implements 18…

计算机视觉与模式识别 · 计算机科学 2022-07-20 Yue Zhou , Xue Yang , Gefan Zhang , Jiabao Wang , Yanyi Liu , Liping Hou , Xue Jiang , Xingzhao Liu , Junchi Yan , Chengqi Lyu , Wenwei Zhang , Kai Chen

Virtual try-on has made significant progress in recent years. This paper addresses how to achieve multifunctional virtual try-on guided solely by text instructions, including full outfit change and local editing. Previous methods primarily…

计算机视觉与模式识别 · 计算机科学 2025-07-09 Yujie Hu , Xuanyu Zhang , Weiqi Li , Jian Zhang

Understanding fashion images has been advanced by benchmarks with rich annotations such as DeepFashion, whose labels include clothing categories, landmarks, and consumer-commercial image pairs. However, DeepFashion has nonnegligible issues…

计算机视觉与模式识别 · 计算机科学 2019-01-24 Yuying Ge , Ruimao Zhang , Lingyun Wu , Xiaogang Wang , Xiaoou Tang , Ping Luo

We present VLMEvalKit: an open-source toolkit for evaluating large multi-modality models based on PyTorch. The toolkit aims to provide a user-friendly and comprehensive framework for researchers and developers to evaluate existing…

We present FairX, an open-source Python-based benchmarking tool designed for the comprehensive analysis of models under the umbrella of fairness, utility, and eXplainability (XAI). FairX enables users to train benchmarking bias-mitigation…

机器学习 · 计算机科学 2024-09-04 Md Fahim Sikder , Resmi Ramachandranpillai , Daniel de Leng , Fredrik Heintz

Personalized outfit recommendation remains a complex challenge, demanding both fashion compatibility understanding and trend awareness. This paper presents a novel framework that harnesses the expressive power of large language models…

信息检索 · 计算机科学 2024-09-19 Najmeh Forouzandehmehr , Nima Farrokhsiar , Ramin Giahi , Evren Korpeoglu , Kannan Achan
‹ 上一页 1 2 3 10 下一页 ›