Related papers: An Implantable Piezofilm Middle Ear Microphone: Pe…
Far-field speech processing is an important and challenging problem. In this paper, we propose \textit{deep ad-hoc beamforming}, a deep-learning-based multichannel speech enhancement framework based on ad-hoc microphone arrays, to address…
In this paper the dielectric properties of human trabecular bone are evaluated under physiological condition in the microwave range. Assuming a two components medium, simulation and experimental data are presented and discussed. A special…
Acoustical knee health assessment has long promised an alternative to clinically available medical imaging tools, but this modality has yet to be adopted in medical practice. The field is currently led by machine learning models processing…
The goal of this work is to train discriminative cross-modal embeddings without access to manually annotated data. Recent advances in self-supervised learning have shown that effective representations can be learnt from natural cross-modal…
Speech embeddings are fixed-size acoustic representations of variable-length speech sequences. They are increasingly used for a variety of tasks ranging from information retrieval to unsupervised term discovery and speech segmentation.…
The high-intensity, repetitive noise associated with functional magnetic resonance imaging hinders on-line monitoring of subjects' speech and/or recording speech signals suitable for off-line analysis. The proposed algorithm enhances the…
Recent Large Audio-Language Models (LALMs) exhibit impressive capabilities in understanding audio content for conversational QA tasks. However, these models struggle to accurately understand timestamps for temporal localization (e.g.,…
The proposed Circular statistics-based Inter-Microphone Phase difference estimation Localizer (CIMPL) method is tailored toward binaural hearing aid systems with microphone arrays in each unit. The method utilizes the circular statistics…
In mobile speech communication applications, wind noise can lead to a severe reduction of speech quality and intelligibility. Since the performance of speech enhancement algorithms using acoustic microphones tends to substantially degrade…
Cochlear Implant (CI) surgery treats severe hearing loss by inserting an electrode array into the cochlea to stimulate the auditory nerve. An important step in this procedure is mastoidectomy, which removes part of the mastoid region of the…
Silent and whispered speech offer promise for always-available voice interaction with AI, yet existing methods struggle to balance vocabulary size, wearability, silence, and noise robustness. We present NasoVoce, a nose-bridge-mounted…
Background: An early diagnosis together with an accurate disease progression monitoring of multiple sclerosis is an important component of successful disease management. Prior studies have established that multiple sclerosis is correlated…
In recent years, remarkable advances in photonic computing have highlighted the need for photonic memory, particularly high-speed and coherent random-access memory. Addressing the ongoing challenge of implementing photonic memories is…
Automatic dietary monitoring has progressed significantly during the last years, offering a variety of solutions, both in terms of sensors and algorithms as well as in terms of what aspect or parameters of eating behavior are measured and…
Sensors and actuators based on resonant micro-electro-mechanical systems (MEMS), such as scanning micro mirrors, are well-established in automotive and consumer products. As the areas of application broaden, the requirements for the MEMS…
The acoustic performance of a new type of fibre-less sound absorber, the micro-grooved element (MGE), is studied in this paper. A MGE is a double layer element that involves inlet/outlet slots and inner micro-channels engraved onto the…
In this work, a novel approach for the detection and localisation of nonlinear guided waves often associated with the presence of damage in structural components is proposed. The method is active and consists of a piezoelectric transducer…
Transcranial focused ultrasound applications often use simulations that require accurate acoustic properties, which can be related to computed tomography (CT) Hounsfield Units (HU). However, clinical CT is insensitive to microstructure.…
Human-imitated speech poses a greater challenge than AI-generated speech for both human listeners and automatic detection systems. Unlike AI-generated speech, which often contains artifacts, over-smoothed spectra, or robotic cues, imitated…
Voice Activity Detection (VAD) is a fundamental module in many audio applications. Recent state-of-the-art VAD systems are often based on neural networks, but they require a computational budget that usually exceeds the capabilities of a…