English
Related papers

Related papers: Long-range temporal correlation in Auditory Brains…

200 papers

Phase aberrations, despite degrading ultrasound images, also encode valuable information about the spatial distribution of the speed of sound in tissue. In pulse-echo ultrasound, we can quantify them by exploiting speckle correlations.…

Medical Physics · Physics 2025-11-20 Naiara Korta Martiartu , Michael Jaeger

While Speech Foundation Models (SFMs) excel in various speech tasks, their performance for low-resource tasks such as child Automatic Speech Recognition (ASR) is hampered by limited pretraining data. To address this, we explore different…

Computation and Language · Computer Science 2025-01-16 Natarajan Balaji Shankar , Zilai Wang , Eray Eren , Abeer Alwan

Speaking Style Recognition (SSR) identifies a speaker's speaking style characteristics from speech. Existing style recognition approaches primarily rely on linguistic information, with limited integration of acoustic information, which…

Sound · Computer Science 2025-10-15 Guojian Li , Qijie Shao , Zhixian Zhao , Shuiyuan Wang , Zhonghua Fu , Lei Xie

The increasing use of children's automatic speech recognition (ASR) systems has spurred research efforts to improve the accuracy of models designed for children's speech in recent years. The current approach utilizes either open-source…

Audio and Speech Processing · Electrical Eng. & Systems 2025-02-13 Vishwanath Pratap Singh , Md. Sahidullah , Tomi Kinnunen

The idea that information-processing systems operate near criticality to enhance computational performance is supported by scaling signatures in brain activity. However, external signals raise the question of whether this behavior is…

Neurons and Cognition · Quantitative Biology 2026-02-10 Rubén Calvo , Carles Martorell , Adrián Roig , Miguel A. Muñoz

Existing retrieval methods in Large Language Models show degradation in accuracy when handling temporally distributed conversations, primarily due to their reliance on simple similarity-based retrieval. Unlike existing memory retrieval…

Computation and Language · Computer Science 2025-07-28 Yuki Hou , Haruki Tamoto , Qinghua Zhao , Homei Miyashita

Wearable devices permit the continuous monitoring of biological processes, such as blood glucose metabolism, and behavior, such as sleep quality and physical activity. The continuous monitoring often occurs in epochs of 60 seconds over…

Methodology · Statistics 2024-04-23 Yuanyuan Luan , Roger S. Zoh , Erjia Cui , Xue Lan , Sneha Jadhav , Carmen D. Tekwe

Multimodal question answering tasks can be used as proxy tasks to study systems that can perceive and reason about the world. Answering questions about different types of input modalities stresses different aspects of reasoning such as…

Computation and Language · Computer Science 2019-11-22 Haytham M. Fayek , Justin Johnson

The audio spectrogram is a time-frequency representation that has been widely used for audio classification. One of the key attributes of the audio spectrogram is the temporal resolution, which depends on the hop size used in the Short-Time…

Sound · Computer Science 2024-01-15 Haohe Liu , Xubo Liu , Qiuqiang Kong , Wenwu Wang , Mark D. Plumbley

Humans understand sentences word-by-word, in the order that they hear them. This incrementality entails resolving temporary ambiguities about syntactic relationships. We investigate how humans process these syntactic ambiguities by…

Computation and Language · Computer Science 2024-06-07 Berta Franzluebbers , Donald Dunagan , Miloš Stanojević , Jan Buys , John T. Hale

Reliable brain tumor segmentation in MRI is indispensable for treatment planning and outcome monitoring, yet models trained on curated benchmarks often fail under domain shifts arising from scanner and protocol variability as well as…

Computer Vision and Pattern Recognition · Computer Science 2025-09-23 Yuanhan Wang , Yifei Chen , Shuo Jiang , Wenjing Yu , Mingxuan Liu , Beining Wu , Jinying Zong , Feiwei Qin , Changmiao Wang , Qiyuan Tian

Correlation-based auditory attention decoding (AAD) algorithms exploit neural tracking mechanisms to determine listener attention among competing speech sources via, e.g., electroencephalography signals. The correlation coefficients between…

Signal Processing · Electrical Eng. & Systems 2025-06-17 Simon Geirnaert , Jonas Vanthornhout , Tom Francart , Alexander Bertrand

The representations generated by many models of language (word embeddings, recurrent neural networks and transformers) correlate to brain activity recorded while people read. However, these decoding results are usually based on the brain's…

Computation and Language · Computer Science 2020-10-16 Maryam Hashemzadeh , Greta Kaufeld , Martha White , Andrea E. Martin , Alona Fyshe

This paper studies the error metric selection for long-term memory learning in sequence modelling. We examine the bias towards short-term memory in commonly used errors, including mean absolute/squared error. Our findings show that all…

Machine Learning · Computer Science 2023-07-24 Shida Wang , Zhanglu Yan

Time-reversal symmetry breaking is a key feature of nearly all natural sounds, caused by the physics of sound production. While attention has been paid to the response of the auditory system to "natural stimuli," very few psychophysical…

Neurons and Cognition · Quantitative Biology 2013-01-04 Jacob N. Oppenheim , Pavel Isakov , Marcelo O. Magnasco

Dysarthric speech recognition (DSR) enhances the accessibility of smart devices for dysarthric speakers with limited mobility. Previously, DSR research was constrained by the fact that existing datasets typically consisted of isolated…

Sound · Computer Science 2025-07-01 Shiyao Wang , Jiaming Zhou , Shiwan Zhao , Yong Qin

Subjective tinnitus (ST) is generally assumed to be a consequence of hearing loss (HL). In animal studies acoustic trauma can lead to behavioral signs of ST, in human studies ST patients without increased hearing thresholds were found to…

Quantitative Methods · Quantitative Biology 2016-03-16 Patrick Krauss , Konstantin Tziridis , Achim Schilling , Claus Metzner , Holger Schulze

Self-supervised learning (SSL) speech models such as wav2vec and HuBERT have demonstrated state-of-the-art performance on automatic speech recognition (ASR) and proved to be extremely useful in low label-resource settings. However, the…

Sound · Computer Science 2023-10-05 Weiwei Lin , Chenhang He , Man-Wai Mak , Youzhi Tu

Brain metabolism is controlled by complex regulation mechanisms. As part of their nature many complex systems show scaling behavior in their timeseries data. Corresponding scaling exponents can sometimes be used to characterize these…

Condensed Matter · Physics 2007-05-23 Stefan Thurner , Christian Windischberger , Ewald Moser , Markus Barth

Speech encoding models use auditory representations to predict how the human brain responds to spoken language stimuli. Most performant encoding models linearly map the hidden states of artificial neural networks to brain data, but this…

Computation and Language · Computer Science 2025-02-14 Nishitha Vattikonda , Aditya R. Vaidya , Richard J. Antonello , Alexander G. Huth
‹ Prev 1 4 5 6 7 8 10 Next ›