English
Related papers

Related papers: Closed-Form Word Error Rate Analysis for Successiv…

200 papers

A method for efficiently successive cancellation (SC) decoding of polar codes with high-dimensional linear binary kernels (HDLBK) is presented and analyzed. We devise a $l$-expressions method which can obtain simplified recursive formulas…

Information Theory · Computer Science 2017-01-18 Zhiliang Huang , Shiyi Zhang , Feiyan Zhang , Chunjiang Duanmu , Ming Chen

Successive cancellation list (SCL) decoders of polar codes excel in practical performance but pose challenges for theoretical analysis. Existing works either limit their scope to erasure channels or address general channels without taking…

Information Theory · Computer Science 2024-05-28 Hsin-Po Wang , Venkatesan Guruswami

The maximum-likelihood (ML) decoder for symbol detection in large multiple-input multiple-output wireless communication systems is typically computationally prohibitive. In this paper, we study a popular and practical alternative, namely…

Signal Processing · Electrical Eng. & Systems 2018-07-04 Christos Thrampoulidis , Weiyu Xu , Babak Hassibi

Speech separation has been successfully applied as a frontend processing module of conversation transcription systems thanks to its ability to handle overlapped speech and its flexibility to combine with downstream tasks such as automatic…

Audio and Speech Processing · Electrical Eng. & Systems 2021-07-06 Jian Wu , Zhuo Chen , Sanyuan Chen , Yu Wu , Takuya Yoshioka , Naoyuki Kanda , Shujie Liu , Jinyu Li

We investigate the effect of speaker localization on the performance of speech recognition systems in a multispeaker, multichannel environment. Given the speaker location information, speech separation is performed in three stages. In the…

Audio and Speech Processing · Electrical Eng. & Systems 2019-10-25 Sunit Sivasankaran , Emmaneul Vincent , Dominique Fohr

Multistatic integrated sensing and communications (ISAC) systems, which use distributed transmitters and receivers, offer enhanced spatial coverage and sensing accuracy compared to stand-alone ISAC configurations. However, these systems…

Signal Processing · Electrical Eng. & Systems 2025-07-29 Taewon Jeong , Lucas Giroto , Umut Utku Erdem , Christian Karle , Jiyeon Choi , Thomas Zwick , Benjamin Nuss

We present a new approach to secure wireless communications using coherent distributed transmission of signals that are spatially decomposed between a two-element distributed antenna array. High-accuracy distributed coordination of…

Signal Processing · Electrical Eng. & Systems 2025-12-17 Anton Schlegel , Jason M/ Merlo , Samuel Wagner , John B. Lancaster , Jeffrey A. Nanzer

State-level minimum Bayes risk (sMBR) training has become the de facto standard for sequence-level training of speech recognition acoustic models. It has an elegant formulation using the expectation semiring, and gives large improvements in…

Computation and Language · Computer Science 2017-06-12 Matt Shannon

Error correction techniques remain effective to refine outputs from automatic speech recognition (ASR) models. Existing end-to-end error correction methods based on an encoder-decoder architecture process all tokens in the decoding phase,…

Computation and Language · Computer Science 2022-08-10 Jingyuan Yang , Rongjun Li , Wei Peng

Automatic Speech Recognition (ASR) transcription errors are commonly assessed using metrics that compare them with a reference transcription, such as Word Error Rate (WER), which measures spelling deviations from the reference, or semantic…

Computation and Language · Computer Science 2025-01-22 Antoine Tholly , Jane Wottawa , Mickael Rouvier , Richard Dufour

This paper proposes a novel joint channel-estimation and source-detection algorithm using successive interference cancellation (SIC)-aided generative score-based diffusion models. Prior work in this area focuses on massive MIMO scenarios,…

Computer Vision and Pattern Recognition · Computer Science 2025-01-22 Sagnik Bhattacharya , Muhammad Ahmed Mohsin , Kamyar Rajabalifardi , John M. Cioffi

Reliable communication over bandlimited and non-linear channels usually requires equalization to simplify receiver processing. Equalizers that perform joint detection and decoding (JDD) achieve the highest information rates but are often…

Information Theory · Computer Science 2024-08-28 Daniel Plabst , Tobias Prinz , Francesca Diedolo , Thomas Wiegart , Georg Böcherer , Norbert Hanik , Gerhard Kramer

The necessity of accurate channel estimation for Successive and Parallel Interference Cancellation is well known. Iterative channel estimation and channel decoding (for instance by means of the Expectation-Maximization algorithm) is…

Information Theory · Computer Science 2016-11-17 Francisco Lazaro Blasco , Francesco Rossetto

Offline handwriting recognition (HWR) has improved significantly with the advent of deep learning architectures in recent years. Nevertheless, it remains a challenging problem and practical applications often rely on post-processing…

Computer Vision and Pattern Recognition · Computer Science 2023-09-20 Andrey Totev , Tomas Ward

In 1995, Best et al. published a formula for the exact bit error probability for Viterbi decoding of the rate R=1/2, memory m=1 (2-state) convolutional encoder with generator matrix G(D)=(1 1+D) when used to communicate over the binary…

Information Theory · Computer Science 2015-03-19 Irina E. Bocharova , Florian Hug , Rolf Johannesson , Boris D. Kudryashov

Since diarization and source separation of meeting data are closely related tasks, we here propose an approach to perform the two objectives jointly. It builds upon the target-speaker voice activity detection (TS-VAD) diarization approach,…

Audio and Speech Processing · Electrical Eng. & Systems 2025-01-23 Christoph Boeddeker , Aswin Shanmugam Subramanian , Gordon Wichern , Reinhold Haeb-Umbach , Jonathan Le Roux

Decoding speaker's intent is a crucial part of spoken language understanding (SLU). The presence of noise or errors in the text transcriptions, in real life scenarios make the task more challenging. In this paper, we address the spoken…

Computation and Language · Computer Science 2019-10-24 Prashanth Gurunath Shivakumar , Mu Yang , Panayiotis Georgiou

The recently proposed VBx diarization method uses a Bayesian hidden Markov model to find speaker clusters in a sequence of x-vectors. In this work we perform an extensive comparison of performance of the VBx diarization with other…

Audio and Speech Processing · Electrical Eng. & Systems 2021-01-01 Federico Landini , Ján Profant , Mireia Diez , Lukáš Burget

Commercial coherent receivers utilize balanced photodetectors (PDs) with high single-port rejection ratio (SPRR) to mitigate the signal-signal beat interference (SSBI) due to the square-law detection process. As the symbol rates of coherent…

Signal Processing · Electrical Eng. & Systems 2022-03-14 Son Thai Le , Vahid Aref , Junho Cho

Source separation is a crucial pre-processing step for various speech processing tasks, such as automatic speech recognition (ASR). Traditionally, the evaluation metrics for speech separation rely on the matched reference audios and…

Audio and Speech Processing · Electrical Eng. & Systems 2025-10-28 Ari Frummer , Helin Wang , Tianyu Cao , Adi Arbel , Yuval Sieradzki , Oren Gal , Jesús Villalba , Thomas Thebaud , Najim Dehak