中文
相关论文

相关论文: Ambisonic Encoding of Signals From Equatorial Micr…

200 篇论文

For a vector random field that is isotropic and mean square continuous on a sphere and stationary on a temporal domain, this paper derives a general form of its covariance matrix function and provides a series representation for the random…

概率论 · 数学 2016-04-26 Chunsheng Ma

Singing voice synthesis (SVS), as a specific task for generating the vocal singing voice from a music score, has drawn much attention in recent years. SVS faces the challenge that the singing has various pronunciation flexibility…

声音 · 计算机科学 2023-03-16 Yuning Wu , Jiatong Shi , Tao Qian , Dongji Gao , Qin Jin

In most automatic speech recognition (ASR) systems, the audio signal is processed to produce a time series of sensor measurements (e.g., filterbank outputs). This time series encodes semantic information in a speaker-dependent way. An…

声音 · 计算机科学 2019-05-10 David N. Levin

This paper reviews pioneering works in microphone array processing and multichannel speech enhancement, highlighting historical achievements, technological evolution, commercialization aspects, and key challenges. It provides valuable…

音频与语音处理 · 电气工程与系统科学 2025-02-14 Gongping Huang , Jesper R. Jensen , Jingdong Chen , Jacob Benesty , Mads G. Christensen , Akihiko Sugiyama , Gary Elko , Tomas Gaensler

The acoustic radiation force produced by ultrasonic waves is the "workhorse" of particle manipulation in acoustofluidics. Nonspherical particles are also subjected to a mean torque known as the acoustic radiation torque. Together they…

流体动力学 · 物理学 2021-08-04 Everton B. Lima , Glauber T. Silva

Beamforming is a signal processing technique. It has been studied in many areas such as radar, sonar, seismology and wireless communications, to name but a few. It can be used for a myriad of purposes, such as detecting the presence of a…

其他计算机科学 · 计算机科学 2012-12-27 Hidri Adel , Meddeb Souad , Abdulqadir Alaqeeli , Amiri Hamid

The performance of voice-controlled systems is usually influenced by accented speech. To make these systems more robust, the frontend accent recognition (AR) technologies have received increased attention in recent years. As accent is a…

音频与语音处理 · 电气工程与系统科学 2021-05-06 Zhan Zhang , Xi Chen , Yuehai Wang , Jianyi Yang

We have developed a sparse mathematical representation of speech that minimizes the number of active model neurons needed to represent typical speech sounds. The model learns several well-known acoustic features of speech such as harmonic…

神经元与认知 · 定量生物学 2012-09-25 Nicole L. Carlson , Vivienne L. Ming , Michael R. DeWeese

Complex-valued sparse coding is a data representation which employs a dictionary of two-dimensional subspaces, while imposing a sparse, factorial prior on complex amplitudes. When trained on a dataset of natural image patches, it learns…

机器学习 · 计算机科学 2014-02-19 Wiktor Mlynarski

Acoustic signal acts as an essential input to many systems. However, the pure acoustic signal is very difficult to extract, especially in noisy environments. Existing beamforming systems are able to extract the signal transmitted from…

声音 · 计算机科学 2022-10-03 Weiguo Wang , Jinming Li , Meng Jin , Yuan He

Spatial information is a critical clue for multi-channel multi-speaker target speech recognition. Most state-of-the-art multi-channel Automatic Speech Recognition (ASR) systems extract spatial features only during the speech separation…

音频与语音处理 · 电气工程与系统科学 2026-01-27 Yiwen Shao , Yong Xu , Sanjeev Khudanpur , Dong Yu

Obtaining constraints from the largest scales of a galaxy survey is challenging due to the survey mask allowing only partial measurement of large angular modes. This scatters information from the harmonic-space 2-point function away from…

宇宙学与河外天体物理 · 物理学 2022-02-16 Henry S. Grasshorn Gebhardt , Olivier Doré

Deep neural network (DNN)-based speech enhancement algorithms in microphone arrays have now proven to be efficient solutions to speech understanding and speech recognition in noisy environments. However, in the context of ad-hoc microphone…

信号处理 · 电气工程与系统科学 2020-11-04 Nicolas Furnon , Romain Serizel , Irina Illina , Slim Essid

The asymptotic form of the plane wave decomposition into spherical waves, which is used to express the scattering amplitude in terms of phase shifts, is incorrect. We explain why and show how to circumvent the mathematical inconsistency.

物理教育 · 物理学 2009-11-10 Radoslaw Maj , Stanislaw Mrowczynski

Purpose: Field monitoring using field probes allows for accurate measurement of magnetic field perturbations, such as from eddy currents, during MRI scanning. However, errors may result when the spatial variation of the fields is not…

The paper explores the problem of \emph{spectral compressed sensing}, which aims to recover a spectrally sparse signal from a small random subset of its $n$ time domain samples. The signal of interest is assumed to be a superposition of $r$…

信息论 · 计算机科学 2015-01-05 Yuxin Chen , Yuejie Chi

The use of spatial information with multiple microphones can improve far-field automatic speech recognition (ASR) accuracy. However, conventional microphone array techniques degrade speech enhancement performance when there is an array…

音频与语音处理 · 电气工程与系统科学 2021-12-23 Kenichi Kumatani , Minhua Wu , Shiva Sundaram , Nikko Strom , Bjorn Hoffmeister

The problem of source localization with ad hoc microphone networks in noisy and reverberant enclosures, given a training set of prerecorded measurements, is addressed in this paper. The training set is assumed to consist of a limited number…

声音 · 计算机科学 2016-10-18 Bracha Laufer-Goldshtein , Ronen Talmon , Sharon Gannot

We present a framework for the optimal filtering of spherical signals contaminated by realizations of an additive, zero-mean, uncorrelated and anisotropic noise process on the sphere. Filtering is performed in the wavelet domain given by…

信号处理 · 电气工程与系统科学 2021-02-09 Adeem Aslam , Zubair Khalid , Jason D. McEwen

This thesis develops a Transformer model based on Whisper, which extracts melodies and chords from music audio and records them into ABC notation. A comprehensive data processing workflow is customized for ABC notation, including data…

声音 · 计算机科学 2024-10-23 Hongyao Zhang , Bohang Sun
‹ 上一页 1 8 9 10 下一页 ›