English
Related papers

Related papers: Investigating differences in lab-quality and remot…

200 papers

Large audio language models are increasingly used for complex audio understanding tasks, but they struggle with temporal tasks that require precise temporal grounding, such as word alignment and speaker diarization. The standard approach,…

Machine Learning · Computer Science 2026-02-12 Joesph An , Phillip Keung , Jiaqi Wang , Orevaoghene Ahia , Noah A. Smith

Digital journaling creates an authenticity gap: users consciously translate raw emotions into text, often sanitizing narratives even in private writing. We formalize this as Cross-Modal Affective Dissonance Detection (CADD), a directional…

Human-Computer Interaction · Computer Science 2026-05-01 Sumin Lee

Audio zooming, a signal processing technique, enables selective focusing and enhancement of sound signals from a specified region, attenuating others. While traditional beamforming and neural beamforming techniques, centered on creating a…

Audio and Speech Processing · Electrical Eng. & Systems 2023-11-23 Meng Yu , Dong Yu

The homogeneity problem for testing if more than two different samples come from the same population is considered for the case of functional data. The methodological results are motivated by the study of homogeneity of electronic devices…

Loudspeaker-based spatial audio reproduction schemes are increasingly used for evaluating hearing aids in complex acoustic conditions. To further establish the feasibility of this approach, this study investigated the interaction between…

Sound · Computer Science 2015-08-04 Giso Grimm , Stephan Ewert , Volker Hohmann

Paralinguistic and non-linguistic aspects of speech strongly influence listener impressions. While most research focuses on absolute impression scoring, this study investigates relative voice impression estimation (RIE), a framework for…

Sound · Computer Science 2026-02-19 Kenichi Fujita , Yusuke Ijima

With the rise of voice-enabled artificial intelligence (AI) systems, quantitative survey researchers have access to a new data-collection mode: AI telephone surveying. By using AI to conduct phone interviews, researchers can scale…

Computation and Language · Computer Science 2025-07-24 Danny D. Leybzon , Shreyas Tirumala , Nishant Jain , Summer Gillen , Michael Jackson , Cameron McPhee , Jennifer Schmidt

Ambient operation poses a challenge to AFM because in contrast to operation in vacuum or liquid environments, the cantilever dynamics change dramatically from oscillating in air to oscillating in a hydration layer when probing the sample.…

Mesoscale and Nanoscale Physics · Physics 2014-02-24 Daniel S. Wastl , Alfred J. Weymouth , Franz J. Giessibl

This study develops an AI-based pose estimation pipeline for quantifying movement kinematics in resistance training. Using videos from Wolf et al. (2025), comprising 303 recordings of 26 participants performing eight upper-body exercises…

Applications · Statistics 2026-03-20 Adam Diamant

Linear systems such as room acoustics and string oscillations may be modeled as the sum of mode responses, each characterized by a frequency, damping and amplitude. Here, we consider finding the mode parameters from impulse response…

Audio and Speech Processing · Electrical Eng. & Systems 2022-02-24 Orchisama Das , Jonathan S. Abel

Precise indoor localization remains a challenging problem for a variety of essential applications. A promising approach to address this problem is to exchange radio signals between mobile agents and static physical anchors (PAs) that bounce…

Signal Processing · Electrical Eng. & Systems 2022-06-22 Erik Leitinger , Bryan Teague , Wenyu Zhang , Mingchao Liang , Florian Meyer

Recent advances in speech synthesis suggest that limitations such as the lossy nature of the amplitude spectrum with minimum phase approximation and the over-smoothing effect in acoustic modeling can be overcome by using advanced machine…

Audio and Speech Processing · Electrical Eng. & Systems 2018-04-10 Xin Wang , Jaime Lorenzo-Trueba , Shinji Takaki , Lauri Juvela , Junichi Yamagishi

Recent multi-modal audio-language models (ALMs) excel at text-audio retrieval but struggle with frame-wise audio understanding. Prior works use temporal-aware labels or unsupervised training to improve frame-wise capabilities, but they…

The remote microphone technique (RMT) is often used in active noise control (ANC) applications to overcome design constraints in microphone placements by estimating the acoustic pressure at inconvenient locations using a pre-calibrated…

Signal Processing · Electrical Eng. & Systems 2023-07-04 Chung Kwan Lai , Bhan Lam , Dongyuan Shi , Woon-Seng Gan

As AI chatbots become ubiquitous, voice interaction presents a compelling way to enable rapid, high-bandwidth communication for both semantic and social signals. This has driven research into Large Audio Models (LAMs) to power voice-native…

Computation and Language · Computer Science 2025-02-25 Minzhi Li , William Barr Held , Michael J Ryan , Kunat Pipatanakul , Potsawee Manakul , Hao Zhu , Diyi Yang

Adaptive gradient methods (AGMs) have become popular in optimizing the nonconvex problems in deep learning area. We revisit AGMs and identify that the adaptive learning rate (A-LR) used by AGMs varies significantly across the dimensions of…

Machine Learning · Computer Science 2019-09-12 Qianqian Tong , Guannan Liang , Jinbo Bi

Atrial fibrillation (AF) is characterized by irregular electrical impulses originating in the atria, which can lead to severe complications and even death. Due to the intermittent nature of the AF, early and timely monitoring of AF is…

Sound · Computer Science 2024-08-12 Xuanyu Liu , Haoxian Liu , Jiao Li , Zongqi Yang , Yi Huang , Jin Zhang

Sound speed heterogeneities can create aberrations in B-mode ultrasound images by inducing tissue-dependent delays and diffractive effects that conventional beamforming does not incorporate. By using the Fourier split-step method to…

Medical Physics · Physics 2026-05-01 Rehman Ali , Trevor M. Mitcham , Marvin M. Doyley , Nebojsa Duric , Jeremy J. Dahl

The frequency-dependent attenuation of broadband acoustics is often confronted in many different areas. However, the related time domain simulation is rarely found in literature due to enormous technical difficulty. The currently popular…

Computational Engineering, Finance, and Science · Computer Science 2007-05-23 W Chen

The photoacoustic signal in a closed T-cell resonator is generated and measured using laser based photoacoustic spectroscopy. The signal is modelled using the amplitude mode expansion method, which is based on eigenmode expansion and…

Instrumentation and Detectors · Physics 2019-03-28 Said El-Busaidy , Bernd Baumann , Marcus Wolff , Lars Duggen , Henry Bruhns
‹ Prev 1 4 5 6 7 8 10 Next ›