中文
相关论文

相关论文: HRTF upsampling with a generative adversarial netw…

200 篇论文

Learning super-resolution (SR) network without the paired low resolution (LR) and high resolution (HR) image is difficult because direct supervision through the corresponding HR counterpart is unavailable. Recently, many real-world SR…

图像与视频处理 · 电气工程与系统科学 2021-09-21 Kwangjin Yoon

In this work, we propose a robust Head-Related Transfer Function (HRTF)-based polynomial beamformer design which accounts for the influence of a humanoid robot's head on the sound field. In addition, it allows for a flexible steering of our…

声音 · 计算机科学 2016-09-09 Hendrik Barfuss , Marcel Mueglich , Walter Kellermann

Neural networks have proven their capabilities by outperforming many other approaches on regression or classification tasks on various kinds of data. Other astonishing results have been achieved using neural nets as data generators,…

计算机视觉与模式识别 · 计算机科学 2018-10-16 Andrej Junginger , Markus Hanselmann , Thilo Strauss , Sebastian Boblest , Jens Buchner , Holger Ulmer

The classification of acoustic environments allows for machines to better understand the auditory world around them. The use of deep learning in order to teach machines to discriminate between different rooms is a new area of research.…

音频与语音处理 · 电气工程与系统科学 2020-12-07 Constantinos Papayiannis , Christine Evers , Patrick A. Naylor

There is a growing demand for high-resolution (HR) medical images in both the clinical and research applications. Image quality is inevitably traded off with the acquisition time for better patient comfort, lower examination costs, dose,…

图像与视频处理 · 电气工程与系统科学 2021-06-07 Kuan Zhang , Haoji Hu , Kenneth Philbrick , Gian Marco Conte , Joseph D. Sobek , Pouria Rouzrokh , Bradley J. Erickson

In this paper, we address the issue of face hallucination. Most current face hallucination methods rely on two-dimensional facial priors to generate high resolution face images from low resolution face images. These methods are only capable…

计算机视觉与模式识别 · 计算机科学 2021-10-06 Shailza Sharma , Abhinav Dhall , Vinay Kumar

High-resolution (HR) magnetic resonance images (MRI) provide detailed anatomical information important for clinical application and quantitative image analysis. However, HR MRI conventionally comes at the cost of longer scan time, smaller…

计算机视觉与模式识别 · 计算机科学 2018-06-12 Yuhua Chen , Feng Shi , Anthony G. Christodoulou , Zhengwei Zhou , Yibin Xie , Debiao Li

We propose an adversarial learning approach for generating multi-turn dialogue responses. Our proposed framework, hredGAN, is based on conditional generative adversarial networks (GANs). The GAN's generator is a modified hierarchical…

计算与语言 · 计算机科学 2019-06-27 Oluwatobi Olabiyi , Alan Salimov , Anish Khazane , Erik T. Mueller

Realistic hyperspectral image (HSI) super-resolution (SR) techniques aim to generate a high-resolution (HR) HSI with higher spectral and spatial fidelity from its low-resolution (LR) counterpart. The generative adversarial network (GAN) has…

图像与视频处理 · 电气工程与系统科学 2022-10-05 Yue Shi , Liangxiu Han , Lianghao Han , Sheng Chang , Tongle Hu , Darren Dancey

In this research, we explore different ways to improve generative adversarial networks for video super-resolution tasks from a base single image super-resolution GAN model. Our primary objective is to identify potential techniques that…

图像与视频处理 · 电气工程与系统科学 2024-06-25 Daniel Wen

This study presents a deep learning-based framework to reconstruct high-resolution turbulent velocity fields from extremely low-resolution data at various Reynolds numbers using the concept of generative adversarial networks (GANs). A…

流体动力学 · 物理学 2022-02-16 Mustafa Z. Yousif , Linqi Yu , Hee-Chang Lim

Very High Spatial Resolution (VHSR) large-scale SAR image databases are still an unresolved issue in the Remote Sensing field. In this work, we propose such a dataset and use it to explore patch-based classification in urban and periurban…

计算机视觉与模式识别 · 计算机科学 2017-11-07 Dimitrios Marmanis , Wei Yao , Fathalrahman Adam , Mihai Datcu , Peter Reinartz , Konrad Schindler , Jan Dirk Wegner , Uwe Stilla

The difficulty in obtaining labeled data relevant to a given task is among the most common and well-known practical obstacles to applying deep learning techniques to new or even slightly modified domains. The data volumes required by the…

计算机视觉与模式识别 · 计算机科学 2019-09-24 Jonathan Howe , Kyle Pula , Aaron A. Reite

Personalized binaural audio reproduction is the basis of realistic spatial localization, sound externalization, and immersive listening, directly shaping user experience and listening effort. This survey reviews recent advances in deep…

音频与语音处理 · 电气工程与系统科学 2025-09-03 Xikun Lu , Yunda Chen , Zehua Chen , Jie Wang , Mingxing Liu , Hongmei Hu , Chengshi Zheng , Stefan Bleeck , Jinqiu Sang

Continuous multimodal representations suitable for multimodal information retrieval are usually obtained with methods that heavily rely on multimodal autoencoders. In video hyperlinking, a task that aims at retrieving video segments, the…

多媒体 · 计算机科学 2017-05-16 Vedran Vukotic , Christian Raymond , Guillaume Gravier

This study introduces an enhanced approach to video super-resolution by extending ordinary Single-Image Super-Resolution (SISR) Super-Resolution Generative Adversarial Network (SRGAN) structure to handle spatio-temporal data. While SRGAN…

计算机视觉与模式识别 · 计算机科学 2025-07-01 Kağan Çetin , Hacer Akça , Ömer Nezih Gerek

4D Flow Magnetic Resonance Imaging (4D Flow MRI) enables non-invasive quantification of blood flow and hemodynamic parameters. However, its clinical application is limited by low spatial resolution and noise, particularly affecting…

In this paper we present SurfaceNet, an approach for estimating spatially-varying bidirectional reflectance distribution function (SVBRDF) material properties from a single image. We pose the problem as an image translation task and propose…

计算机视觉与模式识别 · 计算机科学 2023-07-04 Giuseppe Vecchio , Simone Palazzo , Concetto Spampinato

Generative adversarial networks (GANs) have been recently adopted for super-resolution, an application closely related to what is referred to as "downscaling" in the atmospheric sciences: improving the spatial resolution of low-resolution…

图像与视频处理 · 电气工程与系统科学 2021-11-12 Jussi Leinonen , Daniele Nerini , Alexis Berne

Generative Adversarial Networks (GANs) produce high-quality images but are challenging to train. They need careful regularization, vast amounts of compute, and expensive hyper-parameter sweeps. We make significant headway on these issues by…

计算机视觉与模式识别 · 计算机科学 2021-11-02 Axel Sauer , Kashyap Chitta , Jens Müller , Andreas Geiger