中文
相关论文

相关论文: Computing Optimal Location of Microphone for Impro…

200 篇论文

The auditory system of humanoid robots has gained increased attention in recent years. This system typically acquires the surrounding sound field by means of a microphone array. Signals acquired by the array are then processed using various…

音频与语音处理 · 电气工程与系统科学 2024-01-05 Vladimir Tourbabin , Boaz Rafaely

The ability of robots to estimate their location is crucial for a wide variety of autonomous operations. In settings where GPS is unavailable, measurements of transmissions from fixed beacons provide an effective means of estimating a…

机器人学 · 计算机科学 2017-09-21 Charles Schaff , David Yunis , Ayan Chakrabarti , Matthew R. Walter

Determining the state of a mobile robot is an essential building block of robot navigation systems. In this paper, we address the problem of estimating the robots pose in an indoor environment using 2D LiDAR data and investigate how modern…

机器人学 · 计算机科学 2023-02-06 Haofei Kuang , Xieyuanli Chen , Tiziano Guadagnino , Nicky Zimmerman , Jens Behley , Cyrill Stachniss

In this paper, we study several microphone channel selection and weighting methods for robust automatic speech recognition (ASR) in noisy conditions. For channel selection, we investigate two methods based on the maximum likelihood (ML)…

声音 · 计算机科学 2016-10-04 Zhaofeng Zhang , Xiong Xiao , Longbiao Wang , EngSiong Chng , Haizhou Li

We propose a novel Neural Steering technique that adapts the target area of a spatial-aware multi-microphone sound source separation algorithm during inference without the necessity of retraining the deep neural network (DNN). To achieve…

音频与语音处理 · 电气工程与系统科学 2024-10-23 Martin Strauss , Wolfgang Mack , María Luis Valero , Okan Köpüklü

As humans, we hear sound every second of our life. The sound we hear is often affected by the acoustics of the environment surrounding us. For example, a spacious hall leads to more reverberation. Room Impulse Responses (RIR) are commonly…

人工智能 · 计算机科学 2023-10-10 Yinfeng Yu , Changan Chen , Lele Cao , Fangkai Yang , Fuchun Sun

In this paper, we analyzed how audio-visual speech enhancement can help to perform the ASR task in a cocktail party scenario. Therefore we considered two simple end-to-end LSTM-based models that perform single-channel audio-visual speech…

音频与语音处理 · 电气工程与系统科学 2019-11-28 Luca Pasa , Giovanni Morrone , Leonardo Badino

It is commonly observed that acoustic echoes hurt performance of sound source localization (SSL) methods. We introduce the concept of microphone array augmentation with echoes (MIRAGE) and show how estimation of early-echo characteristics…

音频与语音处理 · 电气工程与系统科学 2019-06-24 Diego Di Carlo , Antoine Deleforge , Nancy Bertin

Loudspeaker-based spatial audio reproduction schemes are increasingly used for evaluating hearing aids in complex acoustic conditions. To further establish the feasibility of this approach, this study investigated the interaction between…

声音 · 计算机科学 2015-08-04 Giso Grimm , Stephan Ewert , Volker Hohmann

The design of a 9-channel microphone system for location recording of mainly atmospheres will be described. The key concept is matching the recording and reproduction angles of the individual sectors. The rig is designed for the AURO-3D…

音频与语音处理 · 电气工程与系统科学 2020-10-13 Florian Camerer

This paper proposes a flexible multichannel speech enhancement system with the main goal of improving robustness of automatic speech recognition (ASR) in noisy conditions. The proposed system combines a flexible neural mask estimator…

音频与语音处理 · 电气工程与系统科学 2024-06-10 Ante Jukić , Jagadeesh Balam , Boris Ginsburg

Sound source localization (SSL) is a critical technology for determining the position of sound sources in complex environments. However, existing methods face challenges such as high computational costs and precise calibration requirements,…

声音 · 计算机科学 2025-05-28 Yiyuan Yang , Shitong Xu , Niki Trigoni , Andrew Markham

Neural sequence-to-sequence systems deliver state-of-the-art performance for automatic speech recognition. When using appropriate modeling units, e.g., byte-pair encoding, these systems are in principle open vocabulary systems. In practice,…

计算与语言 · 计算机科学 2026-03-05 Christian Huber , Alexander Waibel

In this correspondence, we study the secure multiantenna transmission with artificial noise (AN) under imperfect channel state information in the presence of spatially randomly distributed eavesdroppers. We derive the optimal solutions of…

信息论 · 计算机科学 2016-01-07 Tong-Xing Zheng , Hui-Ming Wang

We investigate the problem of designing optimal classifiers in the strategic classification setting, where the classification is part of a game in which players can modify their features to attain a favorable classification outcome (while…

机器学习 · 计算机科学 2020-05-19 Mark Braverman , Sumegha Garg

Conventional speaker localization algorithms, based merely on the received microphone signals, are often sensitive to adverse conditions, such as: high reverberation or low signal to noise ratio (SNR). In some scenarios, e.g. in meeting…

声音 · 计算机科学 2015-08-14 Bracha Laufer-Goldshtein , Ronen Talmon , Sharon Gannot

We investigate mismatched estimation in the context of the distance geometry problem (DGP). In the DGP, for a set of points, we are given noisy measurements of pairwise distances between the points, and our objective is to determine the…

信号处理 · 电气工程与系统科学 2022-06-14 Mahmoud Abdelkhalek , Dror Baron , Chau-Wai Wong

From the eardrum to the auditory cortex, where acoustic stimuli are decoded, there are several stages of auditory processing and transmission where information may potentially get lost. In this paper, we aim at quantifying the information…

信息论 · 计算机科学 2018-05-03 Mohsen Zareian Jahromi , Adel Zahedi , Jesper Jensen , Jan Østergaard

We show that one can reconstruct the shape of a room with planar walls from the first-order echoes received by four non-planar microphones placed on a drone with generic position and orientation. Both the cases where the source is located…

交换代数 · 数学 2020-01-16 Mireille Boutin , Gregor Kemper

Existing research suggests that automatic speech recognition (ASR) models can benefit from additional contexts (e.g., contact lists, user specified vocabulary). Rare words and named entities can be better recognized with contexts. In this…

音频与语音处理 · 电气工程与系统科学 2024-07-16 Ruizhe Huang , Mahsa Yarmohammadi , Sanjeev Khudanpur , Daniel Povey