English
Related papers

Related papers: Physics-Informed Neural Engine Sound Modeling with…

200 papers

This paper presents a physics-informed neural network (PINN) approach for monitoring the health of diesel engines. The aim is to evaluate the engine dynamics, identify unknown parameters in a "mean value" model, and anticipate maintenance…

Machine Learning · Computer Science 2023-08-29 Kamaljyoti Nath , Xuhui Meng , Daniel J Smith , George Em Karniadakis

Current MRI super-resolution (SR) methods only use existing contrasts acquired from typical clinical sequences as input for the neural network (NN). In turbo spin echo sequences (TSE) the sequence parameters can have a strong influence on…

Medical Physics · Physics 2023-05-15 Hoai Nam Dang , Vladimir Golkov , Thomas Wimmer , Daniel Cremers , Andreas Maier , Moritz Zaiss

A deep neural network (DNN)-based model has been developed to predict non-parametric distributions of durations of phonemes in specified phonetic contexts and used to explore which factors influence durations most. Major factors in US…

Sound · Computer Science 2019-09-09 Xizi Wei , Melvyn Hunt , Adrian Skilling

Audio is a fundamental modality for analyzing speech, music, and environmental sounds. Although pretrained audio models have significantly advanced audio understanding, they remain fragile in real-world settings where data distributions…

Sound · Computer Science 2026-02-04 Chang Li , Kanglei Zhou , Liyuan Wang

It is well known that asynchronous impulsive noise is the main source of distortion that drastically affects the power-line communications (PLC) performance. Recently, more realistic models have been proposed in the literature which better…

Information Theory · Computer Science 2015-02-25 Kassim Khalil , Patrick CORLAY , François-Xavier Coudoux , Marc G. Gazalet , Mohamed Gharbi

Decoherence between qubits is a major bottleneck in quantum computations. Decoherence results from intrinsic quantum and thermal fluctuations as well as noise in the external fields that perform the measurement and preparation processes.…

Quantum Physics · Physics 2024-07-17 Ryan T. Grimm , Joel D. Eaves

This study aims at designing an environment-aware text-to-speech (TTS) system that can generate speech to suit specific acoustic environments. It is also motivated by the desire to leverage massive data of speech audio from heterogeneous…

Audio and Speech Processing · Electrical Eng. & Systems 2022-08-09 Daxin Tan , Guangyan Zhang , Tan Lee

Deep learning-based Personal Sound Zones (PSZs) rely on simulated acoustic transfer functions (ATFs) for training, yet idealized point-source models exhibit large sim-to-real gaps. While physically informed components improve…

Audio and Speech Processing · Electrical Eng. & Systems 2026-03-04 Hao Jiang , Edgar Choueiri

We study the phenomenon of nonlinear stochastic resonance (SR) in a complex noisy system formed by a finite number of interacting subunits driven by rectangular pulsed time periodic forces. We find that very large SR gains are obtained for…

Vocoders received renewed attention as main components in statistical parametric text-to-speech (TTS) synthesis and speech transformation systems. Even though there are vocoding techniques give almost accepted synthesized speech, their high…

Sound · Computer Science 2021-06-22 Mohammed Salah Al-Radhi , Tamás Gábor Csapó , Géza Németh

Recent years have seen the rise of statistical program learning based on neural models as an alternative to traditional rule-based systems for programming by example. Rule-based approaches offer correctness guarantees in an unsupervised way…

Machine Learning · Computer Science 2020-06-08 Raphaël Dang-Nhu

Nonlinear Resonant Ultrasound Spectroscopy (NRUS) experiments that rely on repeated sampling of resonance curves are inherently sensitive to measurement protocol due to evolution of material parameters caused by fast and slow dynamic…

Image and Video Processing · Electrical Eng. & Systems 2026-05-01 Jan Kober , Radovan Zeman , Marco Scalerandi

Statistical parametric speech synthesizers have recently shown their ability to produce natural-sounding and flexible voices. Unfortunately the delivered quality suffers from a typical buzziness due to the fact that speech is vocoded. This…

Sound · Computer Science 2020-01-06 Thomas Drugman , Geoffrey Wilfart , Thierry Dutoit

The process of human speech production involves coordinated respiratory action to elicit acoustic speech signals. Typically, speech is produced when air is forced from the lungs and is modulated by the vocal tract, where such actions are…

In this paper we explore the possibility of maximizing the information represented in spectrograms by making the spectrogram basis functions trainable. We experiment with two different tasks, namely keyword spotting (KWS) and automatic…

Sound · Computer Science 2022-04-26 Kwan Yee Heung , Kin Wai Cheuk , Dorien Herremans

Reconstructing unknown external source functions is an important perception capability for a large range of robotics domains including manipulation, aerial, and underwater robotics. In this work, we propose a Physics-Informed Neural Network…

Robotics · Computer Science 2024-11-05 Youngsun Wi , Jayjun Lee , Miquel Oller , Nima Fazeli

In this paper, we first discuss the main types of noise in a typical pump-probe system, and then focus specifically on terahertz time domain spectroscopy (THz-TDS) setups. We then introduce four statistical models for the noisy pulses…

Instrumentation and Detectors · Physics 2017-11-13 M. Skorobogatiy , J. Sadasivan , H. Guerboukha

Pause insertion, also known as phrase break prediction and phrasing, is an essential part of TTS systems because proper pauses with natural duration significantly enhance the rhythm and intelligibility of synthetic speech. However,…

Audio and Speech Processing · Electrical Eng. & Systems 2023-02-28 Dong Yang , Tomoki Koriyama , Yuki Saito , Takaaki Saeki , Detai Xin , Hiroshi Saruwatari

In recent years, Long Short-Term Memory (LSTM) has become a popular choice for speech separation and speech enhancement task. The capability of LSTM network can be enhanced by widening and adding more layers. However, this would introduce…

Sound · Computer Science 2018-12-27 Suman Samui , Indrajit Chakrabarti , Soumya K. Ghosh

Sound field reconstruction refers to the problem of estimating the acoustic pressure field over an arbitrary region of space, using only a limited set of measurements. Physics-informed neural networks have been adopted to solve the problem…

Audio and Speech Processing · Electrical Eng. & Systems 2025-06-05 Stefano Damiano , Toon van Waterschoot
‹ Prev 1 3 4 5 6 7 10 Next ›