English
Related papers

Related papers: GCI detection from raw speech using a fully-convol…

200 papers

The combined electric and acoustic stimulation (EAS) has demonstrated better speech recognition than conventional cochlear implant (CI) and yielded satisfactory performance under quiet conditions. However, when noise signals are involved,…

Existing deep learning-based speech denoising approaches require clean speech signals to be available for training. This paper presents a deep learning-based approach to improve speech denoising in real-world audio environments by not…

Audio and Speech Processing · Electrical Eng. & Systems 2020-02-25 Nasim Alamdari , Arian Azarang , Nasser Kehtarnavaz

This paper addresses the problem of automatic detection of voice pathologies directly from the speech signal. For this, we investigate the use of the glottal source estimation as a means to detect voice disorders. Three sets of features are…

Sound · Computer Science 2020-01-06 Thomas Drugman , Thomas Dubuisson , Thierry Dutoit

Towards developing effective and efficient brain-computer interface (BCI) systems, precise decoding of brain activity measured by electroencephalogram (EEG), is highly demanded. Traditional works classify EEG signals without considering the…

Signal Processing · Electrical Eng. & Systems 2022-09-19 Yimin Hou , Shuyue Jia , Xiangmin Lun , Ziqian Hao , Yan Shi , Yang Li , Rui Zeng , Jinglei Lv

Extreme mass ratio inspirals (EMRIs) are among the most interesting gravitational wave (GW) sources for space-borne GW detectors. However, successful GW data analysis remains challenging due to many issues, ranging from the difficulty of…

High Energy Astrophysical Phenomena · Physics 2022-07-13 Xue-Ting Zhang , Chris Messenger , Natalia Korsakova , Man Leong Chan , Yi-Ming Hu , Jing-dong Zhang

While deep learning has made impressive progress in speech synthesis and voice conversion, the assessment of the synthesized speech is still carried out by human participants. Several recent papers have proposed deep-learning-based…

Audio and Speech Processing · Electrical Eng. & Systems 2020-11-10 Yeunju Choi , Youngmoon Jung , Hoirin Kim

Generating human language through non-invasive brain-computer interfaces (BCIs) has the potential to unlock many applications, such as serving disabled patients and improving communication. Currently, however, generating language via BCIs…

Computation and Language · Computer Science 2025-11-04 Ziyi Ye , Qingyao Ai , Yiqun Liu , Maarten de Rijke , Min Zhang , Christina Lioma , Tuukka Ruotsalo

Practitioners apply neural networks to increasingly complex problems in natural language processing, such as syntactic parsing and semantic role labeling that have rich output structures. Many such structured-prediction problems require…

Computation and Language · Computer Science 2019-04-23 Jay Yoon Lee , Sanket Vaibhav Mehta , Michael Wick , Jean-Baptiste Tristan , Jaime Carbonell

Generative speech enhancement (GSE) models show great promise in producing high-quality clean speech from noisy inputs, enabling applications such as curating noisy text-to-speech (TTS) datasets into high-quality ones. However, GSE models…

Sound · Computer Science 2026-01-21 Kazuki Yamauchi , Masato Murata , Shogo Seki

This paper presents ECGXtract, a deep learning-based approach for interpretable ECG feature extraction, addressing the limitations of traditional signal processing and black-box machine learning methods. In particular, we develop…

Signal Processing · Electrical Eng. & Systems 2025-11-06 Youssif Abuzied , Hassan AbdEltawab , Abdelrhman Gaber , Tamer ElBatt

Ghost imaging (GI) has been paid attention gradually because of its lens-less imaging capability, turbulence-free imaging and high detection sensitivity. However, low image quality and slow imaging speed restrict the application process of…

Image and Video Processing · Electrical Eng. & Systems 2021-04-08 Yuchen He , Sihong Duan , Jianxing Li , Hui Chen , Huaibin Zheng , Jianbin Liu , Shitao Zhu , Zhuo Xu

Real-time noise regression algorithms are crucial for maximizing the science outcomes of the LIGO, Virgo, and KAGRA gravitational-wave detectors. This includes improvements in the detectability, source localization and pre-merger…

Voiced segments of speech are assumed to be composed of non-stationary acoustic objects which can be described as stationary response of a non-stationary fundamental drive (FD) process and which are furthermore suited to reconstruct the…

Sound · Computer Science 2007-05-23 Friedhelm R. Drepper

The aim of the study is to investigate the complex mechanisms of speech perception and ultimately decode the electrical changes in the brain accruing while listening to speech. We attempt to decode heard speech from intracranial…

Human-Computer Interaction · Computer Science 2025-01-28 Milán András Fodor , Tamás Gábor Csapó , Frigyes Viktor Arthur

Classical parametric speech coding techniques provide a compact representation for speech signals. This affords a very low transmission rate but with a reduced perceptual quality of the reconstructed signals. Recently, autoregressive deep…

Audio and Speech Processing · Electrical Eng. & Systems 2019-07-02 Ahmed Mustafa , Arijit Biswas , Christian Bergler , Julia Schottenhamml , Andreas Maier

Expressive speech synthesis aims to generate speech that captures a wide range of para-linguistic features, including emotion and articulation, though current research primarily emphasizes emotional aspects over the nuanced articulatory…

Audio and Speech Processing · Electrical Eng. & Systems 2024-06-24 Zehua Kcriss Li , Meiying Melissa Chen , Yi Zhong , Pinxin Liu , Zhiyao Duan

The great majority of current voice technology applications relies on acoustic features characterizing the vocal tract response, such as the widely used MFCC of LPC parameters. Nonetheless, the airflow passing through the vocal folds, and…

Sound · Computer Science 2020-01-01 Thomas Drugman , Paavo Alku , Abeer Alwan , Bayya Yegnanarayana

As generative artificial intelligence (GAI) models continue to evolve, their generative capabilities are increasingly enhanced and being used extensively in content generation. Beyond this, GAI also excels in data modeling and analysis,…

Signal Processing · Electrical Eng. & Systems 2023-10-31 Jiacheng Wang , Hongyang Du , Dusit Niyato , Jiawen Kang , Shuguang Cui , Xuemin Shen , Ping Zhang

This review summarises the status of silent speech interface (SSI) research. SSIs rely on non-acoustic biosignals generated by the human body during speech production to enable communication whenever normal verbal communication is not…

Audio and Speech Processing · Electrical Eng. & Systems 2020-09-29 Jose A. Gonzalez-Lopez , Alejandro Gomez-Alanis , Juan M. Martín-Doñas , José L. Pérez-Córdoba , Angel M. Gomez

Gait recognition has attracted increasing attention from academia and industry as a human recognition technology from a distance in non-intrusive ways without requiring cooperation. Although advanced methods have achieved impressive success…

Computer Vision and Pattern Recognition · Computer Science 2024-08-14 Guozhen Peng , Yunhong Wang , Yuwei Zhao , Shaoxiong Zhang , Annan Li
‹ Prev 1 3 4 5 6 7 10 Next ›