English
Related papers

Related papers: HiGarment: Cross-modal Harmony Based Diffusion Mod…

200 papers

3D shape generation from text is a fundamental task in 3D representation learning. The text-shape pairs exhibit a hierarchical structure, where a general text like ``chair" covers all 3D shapes of the chair, while more detailed prompts…

Computer Vision and Pattern Recognition · Computer Science 2024-05-01 Zhiying Leng , Tolga Birdal , Xiaohui Liang , Federico Tombari

Diffusion models have become a new generative paradigm for text generation. Considering the discrete categorical nature of text, in this paper, we propose GlyphDiffusion, a novel diffusion approach for text generation via text-guided image…

Computation and Language · Computer Science 2023-05-09 Junyi Li , Wayne Xin Zhao , Jian-Yun Nie , Ji-Rong Wen

The rapid development of generative diffusion models has significantly advanced the field of style transfer. However, most current style transfer methods based on diffusion models typically involve a slow iterative optimization process,…

Computer Vision and Pattern Recognition · Computer Science 2024-10-28 Feihong He , Gang Li , Fuhui Sun , Mengyuan Zhang , Lingyu Si , Xiaoyan Wang , Li Shen

High-fidelity haptic feedback is essential for immersive virtual environments, yet authoring realistic tactile textures remains a significant bottleneck for designers. We introduce HapticMatch, a visual-to-tactile generation framework…

Human-Computer Interaction · Computer Science 2026-01-26 Mingxin Zhang , Yu Yao , Yasutoshi Makino , Hiroyuki Shinoda , Masashi Sugiyama

Clean images are crucial for visual tasks such as small object detection, especially at high resolutions. However, real-world images are often degraded by adverse weather, and weather restoration methods may sacrifice high-frequency details…

Computer Vision and Pattern Recognition · Computer Science 2025-12-12 Wenjie Li , Jinglei Shi , Jin Han , Heng Guo , Zhanyu Ma

Existing handwritten text generation methods primarily focus on isolated words. However, realistic handwritten text demands attention not only to individual words but also to the relationships between them, such as vertical alignment and…

Computer Vision and Pattern Recognition · Computer Science 2025-08-06 Gang Dai , Yifan Zhang , Yutao Qin , Qiangya Guo , Shuangping Huang , Shuicheng Yan

Generating higher-resolution human-centric scenes with details and controls remains a challenge for existing text-to-image diffusion models. This challenge stems from limited training image size, text encoder capacity (limited tokens), and…

Computer Vision and Pattern Recognition · Computer Science 2024-04-09 Gwanghyun Kim , Hayeon Kim , Hoigi Seo , Dong Un Kang , Se Young Chun

Standard clothing asset generation involves restoring forward-facing flat-lay garment images displayed on a clear background by extracting clothing information from diverse real-world contexts, which presents significant challenges due to…

Computer Vision and Pattern Recognition · Computer Science 2025-10-13 Xianfeng Tan , Yuhan Li , Wenxiang Shang , Yubo Wu , Jian Wang , Xuanhong Chen , Yi Zhang , Ran Lin , Bingbing Ni

Despite significant advances in large-scale text-to-image models, achieving hyper-realistic human image generation remains a desirable yet unsolved task. Existing models like Stable Diffusion and DALL-E 2 tend to generate human images with…

Computer Vision and Pattern Recognition · Computer Science 2024-03-18 Xian Liu , Jian Ren , Aliaksandr Siarohin , Ivan Skorokhodov , Yanyu Li , Dahua Lin , Xihui Liu , Ziwei Liu , Sergey Tulyakov

The fashion industry is increasingly leveraging computer vision and deep learning technologies to enhance online shopping experiences and operational efficiencies. In this paper, we address the challenge of generating high-fidelity tiled…

Computer Vision and Pattern Recognition · Computer Science 2025-01-06 Ioannis Xarchakos , Theodoros Koukopoulos

We introduce a general framework for generating diverse visual content, including ambiguous images, panorama images, mesh textures, and Gaussian splat textures, by synchronizing multiple diffusion processes. We present exhaustive…

Computer Vision and Pattern Recognition · Computer Science 2024-11-05 Jaihoon Kim , Juil Koo , Kyeongmin Yeo , Minhyuk Sung

We propose Magic Clothing, a latent diffusion model (LDM)-based network architecture for an unexplored garment-driven image synthesis task. Aiming at generating customized characters wearing the target garments with diverse text prompts,…

Computer Vision and Pattern Recognition · Computer Science 2024-07-25 Weifeng Chen , Tao Gu , Yuhao Xu , Chengcai Chen

Recent advancements in large vision-language models have enabled highly expressive and diverse vector sketch generation. However, state-of-the-art methods rely on a time-consuming optimization process involving repeated feedback from a…

Computer Vision and Pattern Recognition · Computer Science 2025-02-13 Ellie Arar , Yarden Frenkel , Daniel Cohen-Or , Ariel Shamir , Yael Vinker

Diffusion models have been shown to be capable of generating high-quality images, suggesting that they could contain meaningful internal representations. Unfortunately, the feature maps that encode a diffusion model's internal information…

Computer Vision and Pattern Recognition · Computer Science 2024-04-03 Grace Luo , Lisa Dunlap , Dong Huk Park , Aleksander Holynski , Trevor Darrell

Generative models are now widely used by graphic designers and artists. Prior works have shown that these models remember and often replicate content from their training data during generation. Hence as their proliferation increases, it has…

Computer Vision and Pattern Recognition · Computer Science 2024-04-02 Gowthami Somepalli , Anubhav Gupta , Kamal Gupta , Shramay Palta , Micah Goldblum , Jonas Geiping , Abhinav Shrivastava , Tom Goldstein

Scene understanding based on 3D Gaussian Splatting (3DGS) has recently achieved notable advances. Although 3DGS related methods have efficient rendering capabilities, they fail to address the inherent contradiction between the anisotropic…

Computer Vision and Pattern Recognition · Computer Science 2025-05-27 Q. G. Duan , Benyun Zhao , Mingqiao Han Yijun Huang , Ben M. Chen

While recent advances in virtual try-on (VTON) have achieved realistic garment transfer to human subjects, its inverse task, virtual try-off (VTOFF), which aims to reconstruct canonical garment templates from dressed humans, remains…

Computer Vision and Pattern Recognition · Computer Science 2025-08-07 Angang Zhang , Fang Deng , Hao Chen , Zhongjian Chen , Junyan Li

Text-to-image generation has witnessed great progress, especially with the recent advancements in diffusion models. Since texts cannot provide detailed conditions like object appearance, reference images are usually leveraged for the…

Computer Vision and Pattern Recognition · Computer Science 2024-04-09 Zhiqi Huang , Huixin Xiong , Haoyu Wang , Longguang Wang , Zhiheng Li

Diffusion models have demonstrated exceptional capabilities in generating a broad spectrum of visual content, yet their proficiency in rendering text is still limited: they often generate inaccurate characters or words that fail to blend…

Computer Vision and Pattern Recognition · Computer Science 2024-12-03 Jianyi Zhang , Yufan Zhou , Jiuxiang Gu , Curtis Wigington , Tong Yu , Yiran Chen , Tong Sun , Ruiyi Zhang

Generating high-fidelity, 3D-consistent garment textures remains a challenging problem due to the inherent complexities of garment structures and the stringent requirement for detailed, globally consistent texture synthesis. Existing…

Computer Vision and Pattern Recognition · Computer Science 2026-03-10 Jinbo Wu , Xiaobo Gao , Xing Liu , Chen Zhao , Jialun Liu
‹ Prev 1 8 9 10 Next ›