English
Related papers

Related papers: Weighted delay-and-sum beamforming guided by visua…

200 papers

Recently, the end-to-end approach has been successfully applied to multi-speaker speech separation and recognition in both single-channel and multichannel conditions. However, severe performance degradation is still observed in the…

Joint optimization of multi-channel front-end and automatic speech recognition (ASR) has attracted much interest. While promising results have been reported for various tasks, past studies on its meeting transcription application were…

Audio and Speech Processing · Electrical Eng. & Systems 2020-11-30 Xiaofei Wang , Naoyuki Kanda , Yashesh Gaur , Zhuo Chen , Zhong Meng , Takuya Yoshioka

Speech recognition and other natural language tasks have long benefited from voting-based algorithms as a method to aggregate outputs from several systems to achieve a higher accuracy than any of the individual systems. Diarization, the…

Computation and Language · Computer Science 2020-02-06 Andreas Stolcke , Takuya Yoshioka

Ultrasound B-Mode images are created from data obtained from each element in the transducer array in a process called beamforming. The beamforming goal is to enhance signals from specified spatial locations, while reducing signal from all…

Signal Processing · Electrical Eng. & Systems 2020-07-08 Jaime Tierney , Adam Luchies , Christopher Khan , Brett Byram , Matthew Berger

Multi-source localization is an important and challenging technique for multi-talker conversation analysis. This paper proposes a novel supervised learning method using deep neural networks to estimate the direction of arrival (DOA) of all…

Audio and Speech Processing · Electrical Eng. & Systems 2021-11-30 Aswin Shanmugam Subramanian , Chao Weng , Shinji Watanabe , Meng Yu , Dong Yu

We propose a system that gives a mobile robot the ability to separate simultaneous sound sources. A microphone array is used along with a real-time dedicated implementation of Geometric Source Separation and a post-filter that gives us a…

Robotics · Computer Science 2016-03-09 Jean-Marc Valin , Jean Rouat , François Michaud

In this paper, a binaural beamforming algorithm for hearing aid applications is introduced.The beamforming algorithm is designed to be robust to some error in the estimate of the target speaker direction. The algorithm has two main…

Audio and Speech Processing · Electrical Eng. & Systems 2019-11-21 Hala As'ad , Martin Bouchard , Homayoun Kamkar-Parsi

Biological sensory systems are inherently adaptive, filtering out constant stimuli and prioritizing relative changes, likely enhancing computational and metabolic efficiency. Inspired by active sensing behaviors across a wide range of…

Robotics · Computer Science 2026-05-19 Maral Mordad , Kian Behzad , Debojyoti Biswas , Noah J. Cowan , Milad Siami

Millimeter-wave (mmWave) communication is considered as a key enabler of ultra-high data rates in the future cellular and wireless networks. The need for directional communication between base stations (BSs) and users in mmWave systems,…

Signal Processing · Electrical Eng. & Systems 2020-11-04 Sara Khosravi , Hossein S. Ghadikolaei , Marina Petrova

The binaural minimum-variance distortionless-response (BMVDR) beamformer is a well-known noise reduction algorithm that can be steered using the relative transfer function (RTF) vector of the desired speech source. Exploiting the…

Audio and Speech Processing · Electrical Eng. & Systems 2022-11-22 Nico Gößling , Wiebke Middelberg , Simon Doclo

Delay-and-sum (DAS) is the most widespread digital beamformer in high-frame-rate ultrasound imaging. Its implementation is simple and compatible with real-time applications. In this viewpoint article, we describe the fundamentals of DAS…

Signal Processing · Electrical Eng. & Systems 2021-12-17 Vincent Perrot , Maxime Polichetti , François Varray , Damien Garcia

We present a deep neural network-based method to perform high-precision, robust and real-time 6 DOF visual servoing. The paper describes how to create a dataset simulating various perturbations (occlusions and lighting conditions) from a…

Robotics · Computer Science 2017-06-08 Quentin Bateux , Eric Marchand , Jürgen Leitner , Francois Chaumette , Peter Corke

In immersive humanoid robot teleoperation, there are three main shortcomings that can alter the transparency of the visual feedback: the lag between the motion of the operator's and robot's head due to network communication delays or slow…

Linear-array based photoacoustic images are reconstructed using the conventional delay-and-sum (DAS) beamforming method. Although the DAS beamformer is well suited for PA image formation, reconstructed images are often afflicted by noises,…

Medical Physics · Physics 2022-10-05 Sufayan Mulani , Souradip Paul , Mayanglambam Suheshkumar Singh

Blind-audio-source-separation (BASS) techniques, particularly those with low latency, play an important role in a wide range of real-time systems, e.g., hearing aids, in-car hand-free voice communication, real-time human-machine…

Audio and Speech Processing · Electrical Eng. & Systems 2024-06-17 Kaien Mo , Xianrui Wang , Yichen Yang , Shoji Makino , Jingdong Chen

Due to the large bandwidth available, millimeter-Wave (mmWave) bands are considered a viable opportunity to significantly increase the data rate in cellular and wireless networks. Nevertheless, the need for beamforming and directional…

Systems and Control · Electrical Eng. & Systems 2022-05-23 Sara Khosravi , Hossein S. Ghadikolaeiy , Jens Zander , Marina Petrova

Super-resolution ultrasound (SRUS) imaging through localising and tracking sparse microbubbles has been shown to reveal microvascular structure and flow beyond the wave diffraction limit. Most SRUS studies use standard delay and sum (DAS)…

Beamforming is a signal processing technique. It has been studied in many areas such as radar, sonar, seismology and wireless communications, to name but a few. It can be used for a myriad of purposes, such as detecting the presence of a…

Other Computer Science · Computer Science 2012-12-27 Hidri Adel , Meddeb Souad , Abdulqadir Alaqeeli , Amiri Hamid

Training sequences are designed to probe wireless channels in order to obtain channel state information for block-fading channels. Optimal training sounds the channel using orthogonal beamforming vectors to find an estimate that optimizes…

Information Theory · Computer Science 2014-04-04 Andrew J. Duly , Taejoon Kim , David J. Love , James V. Krogmeier

This paper describes a system that generates speaker-annotated transcripts of meetings by using a microphone array and a 360-degree camera. The hallmark of the system is its ability to handle overlapped speech, which has been an unsolved…

‹ Prev 1 3 4 5 6 7 10 Next ›