English
Related papers

Related papers: Sequential Gating Ensemble Network for Noise Robus…

200 papers

In this paper, we address the problem of face hallucination by proposing a novel multi-scale generative adversarial network (GAN) architecture optimized for face verification. First, we propose a multi-scale generator architecture for face…

Computer Vision and Pattern Recognition · Computer Science 2019-09-19 Hadi Kazemi , Fariborz Taherkhani , Nasser M. Nasrabadi

Understanding shadows from a single image spontaneously derives into two types of task in previous studies, containing shadow detection and shadow removal. In this paper, we present a multi-task perspective, which is not embraced by any…

Computer Vision and Pattern Recognition · Computer Science 2017-12-08 Jifeng Wang , Xiang Li , Le Hui , Jian Yang

Blind face restoration (BFR) from severely degraded face images in the wild is a very challenging problem. Due to the high illness of the problem and the complex unknown degradation, directly training a deep neural network (DNN) usually…

Computer Vision and Pattern Recognition · Computer Science 2021-05-14 Tao Yang , Peiran Ren , Xuansong Xie , Lei Zhang

Face super-resolution is a domain-specific image super-resolution, which aims to generate High-Resolution (HR) face images from their Low-Resolution (LR) counterparts. In this paper, we propose a novel face super-resolution method, namely…

Computer Vision and Pattern Recognition · Computer Science 2023-01-04 Xiang Wang , Yimin Yang , Qixiang Pang , Xiao Lu , Yu Liu , Shan Du

In this paper, a neural network named Sequence-to-sequence ConvErsion NeTwork (SCENT) is presented for acoustic modeling in voice conversion. At training stage, a SCENT model is estimated by aligning the feature sequences of source and…

Sound · Computer Science 2020-01-14 Jing-Xuan Zhang , Zhen-Hua Ling , Li-Juan Liu , Yuan Jiang , Li-Rong Dai

Recent years have witnessed promising results of face detection using deep learning. Despite making remarkable progresses, face detection in the wild remains an open research challenge especially when detecting faces at vastly different…

Computer Vision and Pattern Recognition · Computer Science 2018-09-11 Jialiang Zhang , Xiongwei Wu , Jianke Zhu , Steven C. H. Hoi

Generative model based compact video compression is typically operated within a relative narrow range of bitrates, and often with an emphasis on ultra-low rate applications. There has been an increasing consensus in the video communication…

Computer Vision and Pattern Recognition · Computer Science 2025-02-25 Bolin Chen , Hanwei Zhu , Shanzhi Yin , Lingyu Zhu , Jie Chen , Ru-Ling Liao , Shiqi Wang , Yan Ye

Although face recognition has made impressive progress in recent years, we ignore the racial bias of the recognition system when we pursue a high level of accuracy. Previous work found that for different races, face recognition networks…

Computer Vision and Pattern Recognition · Computer Science 2023-04-06 Linzhi Huang , Mei Wang , Jiahao Liang , Weihong Deng , Hongzhi Shi , Dongchao Wen , Yingjie Zhang , Jian Zhao

In this paper, we investigate a deep learning approach for speech denoising through an efficient ensemble of specialist neural networks. By splitting up the speech denoising task into non-overlapping subproblems and introducing a…

Audio and Speech Processing · Electrical Eng. & Systems 2020-08-11 Aswin Sivaraman , Minje Kim

Generative Adversarial Networks (GANs) have the capability of synthesizing images, which have been successfully applied to medical image synthesis tasks. However, most of existing methods merely consider the global contextual information…

Computer Vision and Pattern Recognition · Computer Science 2019-08-14 Tianyang Zhang , Huazhu Fu , Yitian Zhao , Jun Cheng , Mengjie Guo , Zaiwang Gu , Bing Yang , Yuting Xiao , Shenghua Gao , Jiang Liu

Deep learning is a hot research topic in the field of machine learning methods and applications. Generative Adversarial Networks (GANs) and Variational Auto-Encoders (VAEs) provide impressive image generations from Gaussian white noise, but…

Computer Vision and Pattern Recognition · Computer Science 2020-07-29 Jiasong Wu , Jing Zhang , Fuzhi Wu , Youyong Kong , Guanyu Yang , Lotfi Senhadji , Huazhong Shu

We propose spatial semantic embedding network (SSEN), a simple, yet efficient algorithm for 3D instance segmentation using deep metric learning. The raw 3D reconstruction of an indoor environment suffers from occlusions, noise, and is…

Computer Vision and Pattern Recognition · Computer Science 2020-07-08 Dongsu Zhang , Junha Chun , Sang Kyun Cha , Young Min Kim

This study introduces an enhanced approach to video super-resolution by extending ordinary Single-Image Super-Resolution (SISR) Super-Resolution Generative Adversarial Network (SRGAN) structure to handle spatio-temporal data. While SRGAN…

Computer Vision and Pattern Recognition · Computer Science 2025-07-01 Kağan Çetin , Hacer Akça , Ömer Nezih Gerek

In this paper, a novel strategy of Secure Steganograpy based on Generative Adversarial Networks is proposed to generate suitable and secure covers for steganography. The proposed architecture has one generative network, and two…

Computer Vision and Pattern Recognition · Computer Science 2018-11-27 Haichao Shi , Jing Dong , Wei Wang , Yinlong Qian , Xiaoyu Zhang

Generating a pose-invariant representation capable of synthesizing multiple face pose views from a single pose is still a difficult problem. The solution is demanded in various areas like multimedia security, computer vision, robotics, etc.…

Computer Vision and Pattern Recognition · Computer Science 2020-01-06 Hamed Alqahtani

Recent advancement in Generative Adversarial Networks in speech synthesis domain[3],[2] have shown, that it's possible to train GANs [8] in a reliable manner for high quality coherent waveform generation from mel-spectograms. We propose…

Audio and Speech Processing · Electrical Eng. & Systems 2020-06-16 Luka Chkhetiani , Levan Bejanidze

AI-generated images are becoming increasingly realistic and diverse, posing significant challenges for generalizable detection. While Vision Foundation Models (VFMs) provide rich semantic representations and frequency-based methods capture…

Computer Vision and Pattern Recognition · Computer Science 2026-05-01 Shuchang Zhou , Shangkun Wu , Jiwei Wei , Ke Liu , Ran Ran , Caiyan Qin , Yang Yang

Contemporary benchmark methods for image inpainting are based on deep generative models and specifically leverage adversarial loss for yielding realistic reconstructions. However, these models cannot be directly applied on image/video…

Computer Vision and Pattern Recognition · Computer Science 2017-11-20 Avisek Lahiri , Arnav Jain , Prabir Kumar Biswas , Pabitra Mitra

Most of current display devices are with eight or higher bit-depth. However, the quality of most multimedia tools cannot achieve this bit-depth standard for the generating images. De-quantization can improve the visual quality of low…

Image and Video Processing · Electrical Eng. & Systems 2020-04-08 Yang Zhang , Changhui Hu , Xiaobo Lu

This paper describes a general, scalable, end-to-end framework that uses the generative adversarial network (GAN) objective to enable robust speech recognition. Encoders trained with the proposed approach enjoy improved invariance by…

Computation and Language · Computer Science 2017-11-07 Anuroop Sriram , Heewoo Jun , Yashesh Gaur , Sanjeev Satheesh