中文
相关论文

相关论文: Rethinking the Separation Layers in Speech Separat…

200 篇论文

Dual-path is a popular architecture for speech separation models (e.g. Sepformer) which splits long sequences into overlapping chunks for its intra- and inter-blocks that separately model intra-chunk local features and inter-chunk global…

音频与语音处理 · 电气工程与系统科学 2024-03-12 Jia Qi Yip , Shengkui Zhao , Yukun Ma , Chongjia Ni , Chong Zhang , Hao Wang , Trung Hieu Nguyen , Kun Zhou , Dianwen Ng , Eng Siong Chng , Bin Ma

This paper considers multiple-input multiple-output (MIMO) relay communication in multi-cellular (interference) systems in which MIMO source-destination pairs communicate simultaneously. It is assumed that due to severe attenuation and/or…

信息论 · 计算机科学 2017-03-28 Muhammad R A Khandaker , Kai-Kit Wong

Substantial improvement in accuracy of identified linear time-invariant single-input multi-output (SIMO) dynamical models is possible when the disturbances affecting the output measurements are spatially correlated. Using an orthogonal…

系统与控制 · 计算机科学 2015-01-14 Niklas Everitt , Giulio Bottegal , Cristian R. Rojas , Håkan Hjalmarsson

In multiple antenna systems, phase noise due to instabilities of the radio-frequency (RF) oscillators, acts differently depending on whether the RF circuitries connected to each antenna are driven by separate (independent) local oscillators…

信息论 · 计算机科学 2016-11-17 M. Reza Khanzadi , Giuseppe Durisi , Thomas Eriksson

Multiple-input multiple-output has been a key technology for wireless systems for decades. For typical MIMO communication systems, antenna array elements are usually separated by half of the carrier wavelength, thus termed as conventional…

信号处理 · 电气工程与系统科学 2024-08-06 Huizhi Wang , Chao Feng , Yong Zeng , Shi Jin , Chau Yuen , Bruno Clerckx , Rui Zhang

This paper considers the multi-input multi-output (MIMO) relay channel where multiple antennas are employed by each terminal. Compared to single-input single-output (SISO) relay channels, MIMO relay channels introduce additional degrees of…

信息论 · 计算机科学 2008-05-03 Caleb K. Lo , Sriram Vishwanath , Robert W. Heath

Speech separation has been shown effective for multi-talker speech recognition. Under the ad hoc microphone array setup where the array consists of spatially distributed asynchronous microphones, additional challenges must be overcome as…

声音 · 计算机科学 2021-03-04 Dongmei Wang , Takuya Yoshioka , Zhuo Chen , Xiaofei Wang , Tianyan Zhou , Zhong Meng

As the performance of single-channel speech separation systems has improved, there has been a desire to move to more challenging conditions than the clean, near-field speech that initial systems were developed on. When training deep…

音频与语音处理 · 电气工程与系统科学 2021-02-23 Matthew Maciejewski , Jing Shi , Shinji Watanabe , Sanjeev Khudanpur

This work considers communication networks where individual links can be described as MIMO channels. Unlike orthogonal modulation methods (such as the singular-value decomposition), we allow interference between sub-channels, which can be…

信息论 · 计算机科学 2016-11-17 Anatoly Khina , Yuval Kochman , Uri Erez

The wireless communication systems has gone from different generations from SISO systems to MIMO systems. Bandwidth is one important constraint in wireless communication. In wireless communication, high data transmission rates are essential…

网络与互联网体系结构 · 计算机科学 2014-04-01 Kritika Sengar , Nishu Rani , Ankita Singhal , Dolly Sharma , Seema Verma , Tanya Singh

The theory of multiple-input multiple-output (MIMO) technology has been well-developed to increase fading channel capacity over single-input single-output (SISO) systems. This capacity gain can often be leveraged by utilizing channel state…

信息论 · 计算机科学 2008-02-25 Il Han Kim , David J. Love

Speech separation and enhancement (SSE) has advanced remarkably and achieved promising results in controlled settings, such as a fixed number of speakers and a fixed array configuration. Towards a universal SSE system, single-channel…

Speech separation has been successfully applied as a frontend processing module of conversation transcription systems thanks to its ability to handle overlapped speech and its flexibility to combine with downstream tasks such as automatic…

音频与语音处理 · 电气工程与系统科学 2021-07-06 Jian Wu , Zhuo Chen , Sanyuan Chen , Yu Wu , Takuya Yoshioka , Naoyuki Kanda , Shujie Liu , Jinyu Li

Progress in machine learning (ML) has been fueled by scaling neural network models. This scaling has been enabled by ever more heroic feats of engineering, necessary for accommodating ML approaches that require high bandwidth communication…

Speech enhancement and separation have been a long-standing problem, especially with the recent advances using a single microphone. Although microphones perform well in constrained settings, their performance for speech separation decreases…

音频与语音处理 · 电气工程与系统科学 2022-04-15 Muhammed Zahid Ozturk , Chenshu Wu , Beibei Wang , Min Wu , K. J. Ray Liu

We present an upper bound for the Single Channel Speech Separation task, which is based on an assumption regarding the nature of short segments of speech. Using the bound, we are able to show that while the recent methods have made…

音频与语音处理 · 电气工程与系统科学 2023-05-23 Shahar Lutati , Eliya Nachmani , Lior Wolf

Source separation is a fundamental task in speech, music, and audio processing, and it also provides cleaner and larger data for training generative models. However, improving separation performance in practice often depends on increasingly…

声音 · 计算机科学 2025-10-15 Yongsheng Feng , Yuetonghui Xu , Jiehui Luo , Hongjia Liu , Xiaobing Li , Feng Yu , Wei Li

Achieving an increase in the spectral efficiency (SE) has always been a major driver in the design of communication systems. The use of MIMO techniques in mobile communications has achieved significant benefits in improving the system…

信号处理 · 电气工程与系统科学 2018-03-22 Pol Henarejos , Ana I. Pérez-Neira

Speech separation refers to extracting each individual speech source in a given mixed signal. Recent advancements in speech separation and ongoing research in this area, have made these approaches as promising techniques for pre-processing…

机器学习 · 计算机科学 2019-12-18 Fahimeh Bahmaninezhad , Shi-Xiong Zhang , Yong Xu , Meng Yu , John H. L. Hansen , Dong Yu

Speech separation with several speakers is a challenging task because of the non-stationarity of the speech and the strong signal similarity between interferent sources. Current state-of-the-art solutions can separate well the different…

信号处理 · 电气工程与系统科学 2021-02-09 Nicolas Furnon , Romain Serizel , Irina Illina , Slim Essid