English
Related papers

Related papers: Quantifying and Correlating Rhythm Formants in Spe…

200 papers

In this work we formulate a generalized theoretical model to describe the nonlinear dynamics observed in combined frequency-amplitude modulators whose characteristic parameters exhibit a nonlinear dependence on the input modulating signal.…

Mathematical Physics · Physics 2015-05-13 G. Consolo , V. Puliafito , G. Finocchio , L. Lopez-Diaz , R. Zivieri , L. Giovannini , F. Nizzoli , G. Valenti , B. Azzerboni

In this paper, we present a novel multi-modal deep neural network architecture that uses speech and text entanglement for learning phonetically sound spoken-word representations. STEPs-RL is trained in a supervised manner to predict the…

Computation and Language · Computer Science 2020-11-24 Prakamya Mishra

Deep learning approaches have emerged that aim to transform an audio signal so that it sounds as if it was recorded in the same room as a reference recording, with applications both in audio post-production and augmented reality. In this…

Audio and Speech Processing · Electrical Eng. & Systems 2021-07-16 Christian J. Steinmetz , Vamsi Krishna Ithapu , Paul Calamia

In this paper, we introduce new methods and discuss results of text-based LSTM (Long Short-Term Memory) networks for automatic music composition. The proposed network is designed to learn relationships within text documents that represent…

Artificial Intelligence · Computer Science 2016-04-20 Keunwoo Choi , George Fazekas , Mark Sandler

Morphologically rich languages accentuate two properties of distributional vector space models: 1) the difficulty of inducing accurate representations for low-frequency word forms; and 2) insensitivity to distinct lexical relations that…

Computation and Language · Computer Science 2017-06-02 Ivan Vulić , Nikola Mrkšić , Roi Reichart , Diarmuid Ó Séaghdha , Steve Young , Anna Korhonen

Multiple scattering of wave in strong heterogeneity can cause resonance-like wave anomaly where the signal exhibits low-frequency, high intensity, and slowly propagating wave packet velocity. For example, long period event in volcanic…

Classical Physics · Physics 2019-12-19 Yinbin Liu

In general, multi-channel source separation has utilized inter-microphone phase differences (IPDs) concatenated with magnitude information in time-frequency domain, or real and imaginary components stacked along the channel axis. However,…

Audio and Speech Processing · Electrical Eng. & Systems 2026-04-01 Ui-Hyeop Shin , Bon Hyeok Ku , Hyung-Min Park

In this study, we employ the atomistic wave-packet method to directly simulate coherent phonon transport and scattering dynamics in an aperiodic superlattice structure with aperiodically arranged interfaces. Our investigation reveals that…

Materials Science · Physics 2024-09-04 Theodore Maranets , Milad Nasiri , Yan Wang

Identification of the type of communication technology and/or modulation scheme based on detected radio signal are challenging problems encountered in a variety of applications including spectrum allocation and radio interference…

Signal Processing · Electrical Eng. & Systems 2020-11-18 Ziqi Ke , Haris Vikalo

Nanomechanical resonators promise diverse applications ranging from mass spectrometry to quantum information processing, requiring long phonon lifetimes and frequency stability. Although two-level system (TLS) defects govern dissipation at…

Mesoscale and Nanoscale Physics · Physics 2025-01-15 M. P. Maksymowych , M. Yuksel , O. A. Hitchcock , N. R. Lee , F. M. Mayor , W. Jiang , M. L. Roukes , A. H. Safavi-Naeini

This study characterises the radio luminosity functions (RLFs) for SFGs and AGN using statistical redshift estimation in the absence of comprehensive spectroscopic data. Sensitive radio surveys over large areas detect many sources with…

Language models (LMs) are being scaled and becoming powerful. Improving their efficiency is one of the core research topics in neural information processing systems. Tay et al. (2022) provided a comprehensive overview of efficient…

Machine Learning · Computer Science 2023-06-06 Meng Jiang , Hy Dang , Lingbo Tong

The Generation and propagation of the human voice is studied in two-dimensions using a full-body domain, using direct numerical simulation. The fluid/air in the vocal tract is modeled as a compressible and viscous fluid interacting with the…

Fluid Dynamics · Physics 2020-05-06 Shakti Saurabh , Daniel Bodony

Noise-robust automatic speech recognition (ASR) has been commonly addressed by applying speech enhancement (SE) at the waveform level before recognition. However, speech-level enhancement does not always translate into consistent…

Audio and Speech Processing · Electrical Eng. & Systems 2026-01-09 Da-Hee Yang , Joon-Hyuk Chang

This paper presents a simple Fourier-matching method to rigorously study resonance frequencies of a sound-hard slab with a finite number of arbitrarily shaped cylindrical holes of diameter ${\cal O}(h)$ for $h\ll1$. Outside the holes, a…

Analysis of PDEs · Mathematics 2021-04-07 Wangtao Lu , Wei Wang , Jiaxin Zhou

Resonant mode interactions in weakly nonlinear multi-dimensional lattices and related effects are described. We concentrate on formal description of the phenomenon and consider as examples mode interactions and evolution equations for…

Statistical Mechanics · Physics 2007-05-23 V. v. Konotop

On the basis of the f-deformed oscillator formalism, we propose to construct nonlinear coherent states for Hamiltonian systems having linear and quadratic terms in the the number operator by means of the two following definitions: i) as…

Quantum Physics · Physics 2015-12-03 R. Román-Ancheyta , J. Récamier

Large Language Models (LLMs) have emerged as powerful support tools across various natural language tasks and a range of application domains. Recent studies focus on exploring their capabilities for data annotation. This paper provides a…

Computation and Language · Computer Science 2025-07-01 Maja Pavlovic , Massimo Poesio

This paper explores the potential of large language models (LLMs) as reliable analytical tools in linguistic research, focusing on the emergence of affective meanings in temporal expressions involving manner-of-motion verbs. While LLMs like…

Computation and Language · Computer Science 2025-07-15 Rosa Illan Castillo , Javier Valenzuela

Recent advances in speech foundation models (SFMs) have enabled the direct processing of spoken language from raw audio, bypassing intermediate textual representations. This capability allows SFMs to be exposed to, and potentially respond…

Audio and Speech Processing · Electrical Eng. & Systems 2025-10-30 Harm Lameris , Shree Harsha Bokkahalli Satish , Joakim Gustafson , Éva Székely
‹ Prev 1 8 9 10 Next ›