中文
相关论文

相关论文: Raking the Cocktail Party

200 篇论文

Radio speech echo is a specific phenomenon in the air traffic control (ATC) domain, which degrades speech quality and further impacts automatic speech recognition (ASR) accuracy. In this work, a time-domain recognition-oriented speech…

声音 · 计算机科学 2024-07-31 Xincheng Yu , Dongyue Guo , Jianwei Zhang , Yi Lin

We consider that a transmitter covertly communicates with multiple receivers under the help of a friendly jammer. The messages intended for different receivers are transmitted in mutually orthogonal frequency bands. An adversary observes…

信息论 · 计算机科学 2021-03-03 Ke-Wen Huang , Hao Deng , Hui-Ming Wang

Hearing aids use dynamic range compression (DRC), a form of automatic gain control, to make quiet sounds louder and loud sounds quieter. Compression can improve listening comfort, but it can also cause distortion in noisy environments. It…

音频与语音处理 · 电气工程与系统科学 2021-07-28 Ryan M. Corey , Andrew C. Singer

Acoustic echo cancellation (AEC) is an important speech signal processing technology that can remove echoes from microphone signals to enable natural-sounding full-duplex speech communication. While single-channel AEC is widely adopted,…

声音 · 计算机科学 2025-06-09 Fei Zhao , Xueliang Zhang , Zhong-Qiu Wang

In this paper, a general cognitive radio system consisting of a set of users with different level of spectrum access including two primary transceivers and several types of secondary users is considered. It is assumed that two secondary…

信息论 · 计算机科学 2018-09-10 Mohammad Zaeri-Amirani , Fatemeh Afghah , Jonathan Ashdown

Covert communication is often limited in rate because it is difficult to hide the signal in the background noise. Recent work has shown that jamming can significantly improve the rate at which covert communications can be conducted;…

信号处理 · 电气工程与系统科学 2019-07-04 Moslem Forouzesh , Paeiz Azmi , Nader Mokari , Dennis Goeckel

Through spatial multiplexing and diversity, multi-input multi-output (MIMO) cognitive radio (CR) networks can markedly increase transmission rates and reliability, while controlling the interference inflicted to peer nodes and primary users…

信息论 · 计算机科学 2013-02-07 Yu Zhang , Emiliano Dall'Anese , Georgios B. Giannakis

To improve signal-to-interference ratio (SIR) and make better use of file diversity provided by random caching, we consider two types of linear receivers, i.e., maximal ratio combining (MRC) receiver and partial zero forcing (PZF) receiver,…

信息论 · 计算机科学 2018-01-10 Dongdong Jiang , Ying Cui

We consider the joint design of transmit beamforming and receive signal-splitting ratios in the downlink of a wireless network with simultaneous radio-frequency (RF) information and energy transfer. Under constraints on the…

信息论 · 计算机科学 2017-10-13 Ali A. Nasir , Hoang D. Tuan , Duy T. Ngo , Salman Durrani , Dong In Kim

Reconfigurable intelligent surface (RIS) technology is a promising solution to improve the performance of existing wireless communications. To achieve its cost-effectiveness advantage, there inevitably exist certain hardware impairments in…

信号处理 · 电气工程与系统科学 2024-01-17 Yiming Liu , Rui Wang , Zhu Han

This paper maximizes the achievable throughput of a relay-assisted wirelessly powered communications system, where an energy constrained source, helped by an energy constrained relay and both powered by a dedicated power beacon (PB),…

信息论 · 计算机科学 2016-11-17 Caijun Zhong , Gan Zheng , Zhaoyang Zhang , George K. Karagiannidis

We propose, analyze and demonstrate an architecture for scalable cooperative reception. In a cluster of N + 1 receive nodes, one node is designated as the final receiver, and the N other nodes act as amplify-and-forward relays which adapt…

信息论 · 计算机科学 2016-11-17 Francois Quitin , Andrew T. Irish , Upamanyu Madhow

Speech enhancement plays an essential role in improving the quality of speech signals in noisy environments. This paper investigates the efficacy of integrating Bidirectional Gated Recurrent Units (BGRU) and Transformer models for speech…

声音 · 计算机科学 2025-02-26 Souliman Alghnam , Mohammad Alhussien , Khaled Shaheen

This paper describes a framework and a method with which speech communication can be analyzed. The framework consists of a set of low bit rate, short-range acoustic communication systems, such as speech, but that are quite different from…

多媒体 · 计算机科学 2010-10-20 Cristina Videira Lopes , Pedro M. Q. Aguiar

Integrated sensing and communication (ISAC), which allows individual radar and communication systems to share the same spectrum bands, is an emerging and promising technique for alleviating spectrum congestion problems. In this paper, we…

信号处理 · 电气工程与系统科学 2022-12-02 Jinjin Chu , Rang Liu , Ming Li , Yang Liu , Qian Liu

This paper presents an efficient speech enhancement (SE) approach that reuses a processing block repeatedly instead of conventional stacking. Rather than increasing the number of blocks for learning deep latent representations, repeating a…

音频与语音处理 · 电气工程与系统科学 2026-04-01 Jangyeon Kim , Ui-Hyeop Shin , Jaehyun Ko , Hyung-Min Park

Deep generative models applied to audio have improved by a large margin the state-of-the-art in many speech and music related tasks. However, as raw waveform modelling remains an inherently difficult task, audio generative models are either…

机器学习 · 计算机科学 2021-12-16 Antoine Caillon , Philippe Esling

Speech separation refers to extracting each individual speech source in a given mixed signal. Recent advancements in speech separation and ongoing research in this area, have made these approaches as promising techniques for pre-processing…

机器学习 · 计算机科学 2019-12-18 Fahimeh Bahmaninezhad , Shi-Xiong Zhang , Yong Xu , Meng Yu , John H. L. Hansen , Dong Yu

This paper investigates the performance of one- and two-sided amplitude shift keying (ASK) modulations in noncoherent single-input single-output (SISO) wireless communication systems assisted by a reconfigurable intelligent surface (RIS).…

信息论 · 计算机科学 2024-12-24 Sambit Mishra , Soumya P. Dash , George C. Alexandropoulos

Diffusion models have achieved state-of-the-art synthesis quality on both visual and audio tasks, and recent works further adapt them to textual data by diffusing on the embedding space. In this paper, we conduct systematic studies of the…

计算与语言 · 计算机科学 2024-04-23 Zhujin Gao , Junliang Guo , Xu Tan , Yongxin Zhu , Fang Zhang , Jiang Bian , Linli Xu