中文
相关论文

相关论文: DTT-BSR: GAN-based DTTNet with RoPE Transformer En…

200 篇论文

The Inaugural Music Source Restoration (MSR) Challenge targets the recovery of original, unprocessed stems from fully mixed and mastered music. Unlike conventional music source separation, MSR requires reversing complex production processes…

声音 · 计算机科学 2026-03-19 Xinlong Deng , Yu Xia , Jie Jiang

Music source separation (MSS) aims to separate a music recording into multiple musically distinct stems, such as vocals, bass, drums, and more. Recently, deep learning approaches such as convolutional neural networks (CNNs) and recurrent…

声音 · 计算机科学 2023-09-12 Wei-Tsung Lu , Ju-Chiang Wang , Qiuqiang Kong , Yun-Ning Hung

Single image super-resolution (SISR) reconstruction for magnetic resonance imaging (MRI) has generated significant interest because of its potential to not only speed up imaging but to improve quantitative processing and analysis of…

图像与视频处理 · 电气工程与系统科学 2019-07-17 Jiancong Wang , Yuhua Chen , Yifan Wu , Jianbo Shi , James Gee

Music Source Restoration (MSR) targets recovery of original, unprocessed instrument stems from fully mixed and mastered audio, where production effects and distribution artifacts violate common linear-mixture assumptions. This technical…

声音 · 计算机科学 2026-03-05 Tobias Morocutti , Emmanouil Karystinaios , Jonathan Greif , Gerhard Widmer

Objective: Many studies on radar signal restoration in the literature focus on isolated restoration problems, such as denoising over a certain type of noise, while ignoring other types of artifacts. Additionally, these approaches usually…

机器学习 · 计算机科学 2024-12-11 Muhammad Uzair Zahid , Serkan Kiranyaz , Alper Yildirim , Moncef Gabbouj

Multi-segment reconstruction (MSR) is the problem of estimating a signal given noisy partial observations. Here each observation corresponds to a randomly located segment of the signal. While previous works address this problem using…

信号处理 · 电气工程与系统科学 2021-02-19 Mona Zehni , Zhizhen Zhao

Speech super-resolution (SSR) aims to recover a high resolution (HR) speech from its corresponding low resolution (LR) counterpart. Recent SSR methods focus more on the reconstruction of the magnitude spectrogram, ignoring the importance of…

音频与语音处理 · 电气工程与系统科学 2023-05-22 Chenhao Shuai , Chaohua Shi , Lu Gan , Hongqing Liu

Recently, deep neural networks (DNNs) have been used extensively for automatic modulation classification (AMC), and the results have been quite promising. However, DNNs have high memory and computation requirements making them impractical…

Music Source Restoration (MSR) extends source separation to realistic settings where signals undergo production effects (equalization, compression, reverb) and real-world degradations, with the goal of recovering the original unprocessed…

Besides the well-known classification task, these days neural networks are frequently being applied to generate or transform data, such as images and audio signals. In such tasks, the conventional loss functions like the mean squared error…

The performance of music source separation (MSS) models has been greatly improved in recent years thanks to the development of novel neural network architectures and training pipelines. However, recent model designs for MSS were mainly…

音频与语音处理 · 电气工程与系统科学 2022-10-03 Yi Luo , Jianwei Yu

High-resolution (HR) magnetic resonance images (MRI) provide detailed anatomical information important for clinical application and quantitative image analysis. However, HR MRI conventionally comes at the cost of longer scan time, smaller…

计算机视觉与模式识别 · 计算机科学 2018-06-12 Yuhua Chen , Feng Shi , Anthony G. Christodoulou , Zhengwei Zhou , Yibin Xie , Debiao Li

Over the past few years, there has been growing interest in developing larger and deeper neural networks, including deep generative models like generative adversarial networks (GANs). However, GANs typically come with high computational…

机器学习 · 计算机科学 2023-11-21 Yite Wang , Jing Wu , Naira Hovakimyan , Ruoyu Sun

We study generative super-resolution (SR) in real-world scenarios where content and degradations vary across domains, genres, and segments. For example, images and videos may alternate between text overlays, fast motion, smooth cartoons,…

计算机视觉与模式识别 · 计算机科学 2026-04-30 Jiaqi Guo , Mingzhen Li , Haohong Wang , Aggelos K. Katsaggelos

Music Source Restoration (MSR) aims to recover original, unprocessed instrument stems from professionally mixed and degraded audio, requiring the reversal of both production effects and real-world degradations. We present the inaugural MSR…

Accelerated Cardiovascular Magnetic Resonance (CMR) image reconstruction remains a critical challenge due to the trade-off between scan time and image quality, particularly when generalizing across diverse acquisition settings. We propose…

图像与视频处理 · 电气工程与系统科学 2025-10-30 Kian Anvari Hamedani , Narges Razizadeh , Shahabedin Nabavi , Mohsen Ebrahimi Moghaddam

Audio super-resolution (SR), also referred to as bandwidth extension (BWE), aims to reconstruct high-fidelity signals from low-resolution (LR) or band-limited (BL) observations, an inherently ill-posed task due to the ambiguity of missing…

音频与语音处理 · 电气工程与系统科学 2026-05-20 Ningyuan Yang , Yize Li , Diego A. Cuji , Ryan M. Corey , Pu Zhao , Xue Lin , Andrew C. Singer

We investigate the use of generative adversarial networks (GANs) in speech dereverberation for robust speech recognition. GANs have been recently studied for speech enhancement to remove additive noises, but there still lacks of a work to…

声音 · 计算机科学 2019-01-01 Ke Wang , Junbo Zhang , Sining Sun , Yujun Wang , Fei Xiang , Lei Xie

Video super-resolution (VSR) has become one of the most critical problems in video processing. In the deep learning literature, recent works have shown the benefits of using adversarial-based and perceptual losses to improve the performance…

计算机视觉与模式识别 · 计算机科学 2019-06-26 Alice Lucas , Santiago Lopez Tapia , Rafael Molina , Aggelos K. Katsaggelos

Recently, video super resolution (VSR) has become a very impactful task in the area of Computer Vision due to its various applications. In this paper, we propose Recurrent Back-Projection Generative Adversarial Network (RBPGAN) for VSR in…

计算机视觉与模式识别 · 计算机科学 2025-01-06 Marwah Sulaiman , Zahraa Shehabeldin , Israa Fahmy , Mohammed Barakat , Mohammed El-Naggar , Dareen Hussein , Moustafa Youssef , Hesham M. Eraqi
‹ 上一页 1 2 3 10 下一页 ›