English
Related papers

Related papers: SE-GAN: Skeleton Enhanced GAN-based Model for Brus…

200 papers

In medical image synthesis, model training could be challenging due to the inconsistencies between images of different modalities even with the same patient, typically caused by internal status/tissue changes as different modalities are…

Image and Video Processing · Electrical Eng. & Systems 2021-09-16 Hajar Emami , Ming Dong , Siamak Nejad-Davarani , Carri Glide-Hurst

The automatic generation of Chinese fonts is an important problem involved in many applications. The predominated methods for the Chinese font generation are based on the deep generative models, especially the generative adversarial…

Computer Vision and Pattern Recognition · Computer Science 2022-11-29 Jie Zhou , Yefei Wang , Yiyang Yuan , Qing Huang , Jinshan Zeng

Recent works have shown Generative Adversarial Networks (GANs) to be particularly effective in image-to-image translations. However, in tasks such as body pose and hand gesture translation, existing methods usually require precise…

Computer Vision and Pattern Recognition · Computer Science 2019-11-11 Yahui Liu , Marco De Nadai , Gloria Zen , Nicu Sebe , Bruno Lepri

Generating and manipulating human facial images using high-level attributal controls are important and interesting problems. The models proposed in previous work can solve one of these two problems (generation or manipulation), but not both…

Computer Vision and Pattern Recognition · Computer Science 2017-04-10 Weidong Yin , Yanwei Fu , Leonid Sigal , Xiangyang Xue

In this paper, we propose a novel cross-attention-based generative adversarial network (GAN) for the challenging person image generation task. Cross-attention is a novel and intuitive multi-modal fusion method in which an…

Computer Vision and Pattern Recognition · Computer Science 2025-01-16 Hao Tang , Ling Shao , Nicu Sebe , Luc Van Gool

Generating realistic biometric images has been an interesting and, at the same time, challenging problem. Classical statistical models fail to generate realistic-looking fingerprint images, as they are not powerful enough to capture the…

Computer Vision and Pattern Recognition · Computer Science 2019-01-09 Shervin Minaee , Amirali Abdolrashidi

Cross-view image translation is challenging because it involves images with drastically different views and severe deformation. In this paper, we propose a novel approach named Multi-Channel Attention SelectionGAN (SelectionGAN) that makes…

Computer Vision and Pattern Recognition · Computer Science 2019-04-18 Hao Tang , Dan Xu , Nicu Sebe , Yanzhi Wang , Jason J. Corso , Yan Yan

Synthesising a text-to-image model of high-quality images by guiding the generative model through the Text description is an innovative and challenging task. In recent years, AttnGAN based on the Attention mechanism to guide GAN training…

Computer Vision and Pattern Recognition · Computer Science 2023-07-07 Mingyu Jin , Chong Zhang , Qinkai Yu , Haochen Xue , Xiaobo Jin , Xi Yang

We present a new perspective of achieving image synthesis by viewing this task as a visual token generation problem. Different from existing paradigms that directly synthesize a full image from a single input (e.g., a latent code), the new…

Computer Vision and Pattern Recognition · Computer Science 2021-12-21 Yanhong Zeng , Huan Yang , Hongyang Chao , Jianbo Wang , Jianlong Fu

The generation of Chinese fonts has a wide range of applications. The currently predominated methods are mainly based on deep generative models, especially the generative adversarial networks (GANs). However, existing GAN-based models…

Computer Vision and Pattern Recognition · Computer Science 2022-11-14 Jinshan Zeng , Yefei Wang , Qi Chen , Yunxin Liu , Mingwen Wang , Yuan Yao

Recent advances in Generative Adversarial Networks (GANs) have shown increasing success in generating photorealistic images. But they also raise challenges to visual forensics and model attribution. We present the first study of learning…

Computer Vision and Pattern Recognition · Computer Science 2019-08-19 Ning Yu , Larry Davis , Mario Fritz

In recent years generative models of visual data have made a great progress, and now they are able to produce images of high quality and diversity. In this work we study representations learnt by a GAN generator. First, we show that these…

Computer Vision and Pattern Recognition · Computer Science 2020-06-19 Danil Galeev , Konstantin Sofiiuk , Danila Rukhovich , Mikhail Romanov , Olga Barinova , Anton Konushin

This paper addresses the problem of diversity-aware sign language production, where we want to give an image (or sequence) of a signer and produce another image with the same pose but different attributes (\textit{e.g.} gender, skin color).…

Computer Vision and Pattern Recognition · Computer Science 2024-05-20 Mohamed Ilyes Lakhal , Richard Bowden

We introduce BSD-GAN, a novel multi-branch and scale-disentangled training method which enables unconditional Generative Adversarial Networks (GANs) to learn image representations at multiple scales, benefiting a wide range of generation…

Computer Vision and Pattern Recognition · Computer Science 2020-08-05 Zili Yi , Zhiqin Chen , Hao Cai , Wendong Mao , Minglun Gong , Hao Zhang

Sign language recognition (SLR) refers to interpreting sign language glosses from given videos automatically. This research area presents a complex challenge in computer vision because of the rapid and intricate movements inherent in sign…

Computer Vision and Pattern Recognition · Computer Science 2025-03-27 Muxin Pu , Mei Kuan Lim , Chun Yong Chong

Electromyography (EMG)-based gesture recognition has emerged as a promising approach for human-computer interaction. However, its performance is often limited by the scarcity of labeled EMG data, significant cross-user variability, and poor…

Human-Computer Interaction · Computer Science 2025-12-11 Nana Wang , Gen Li , Pengfei Ren , Hao Su , Suli Wang

In this paper, we propose a skeleton matching based approach which aids in text localization in scene images. The input image is preprocessed and segmented into blocks using connected component analysis. We obtain the skeleton of the…

Computer Vision and Pattern Recognition · Computer Science 2015-02-23 B. H. Shekar , Smitha M. L

Sign language is commonly used by deaf or speech impaired people to communicate but requires significant effort to master. Sign Language Recognition (SLR) aims to bridge the gap between sign language users and others by recognizing signs…

Computer Vision and Pattern Recognition · Computer Science 2021-05-04 Songyao Jiang , Bin Sun , Lichen Wang , Yue Bai , Kunpeng Li , Yun Fu

In recent years, the use of deep learning is becoming increasingly popular in computer vision. However, the effective training of deep architectures usually relies on huge sets of annotated data. This is critical in the medical field where…

Image and Video Processing · Electrical Eng. & Systems 2019-07-30 Paolo Andreini , Simone Bonechi , Monica Bianchini , Alessandro Mecocci , Franco Scarselli , Andrea Sodi

In this paper, we investigate the Chinese calligraphy synthesis problem: synthesizing Chinese calligraphy images with specified style from standard font(eg. Hei font) images (Fig. 1(a)). Recent works mostly follow the stroke extraction and…

Computer Vision and Pattern Recognition · Computer Science 2017-06-28 Pengyuan Lyu , Xiang Bai , Cong Yao , Zhen Zhu , Tengteng Huang , Wenyu Liu