中文
相关论文

相关论文: A Dynamic Approach to Rhythm in Language: Toward a…

200 篇论文

Converting input symbols to output audio in TTS requires modelling the durations of speech sounds. Leading non-autoregressive (NAR) TTS models treat duration modelling as a regression problem. The same utterance is then spoken with…

音频与语音处理 · 电气工程与系统科学 2024-06-11 Shivam Mehta , Harm Lameris , Rajiv Punmiya , Jonas Beskow , Éva Székely , Gustav Eje Henter

We present DYNARTmo, a dynamic articulatory model designed to visualize speech articulation processes in a two-dimensional midsagittal plane. The model builds upon the UK-DYNAMO framework and integrates principles of articulatory…

计算与语言 · 计算机科学 2025-11-07 Bernd J. Kröger

A hallmark of human language is the ability to effectively and efficiently convey contextually relevant information. One theory for how humans reason about language is presented in the Rational Speech Acts (RSA) framework, which captures…

计算与语言 · 计算机科学 2020-06-02 Julia White , Jesse Mu , Noah D. Goodman

Recent progress in Spoken Language Modeling has shown that learning language directly from speech is feasible. Generating speech through a pipeline that operates at the text level typically loses nuances, intonations, and non-verbal…

计算与语言 · 计算机科学 2024-10-31 Maxime Poli , Emmanuel Chemla , Emmanuel Dupoux

We introduce RHYTHM (Reasoning with Hierarchical Temporal Tokenization for Human Mobility), a framework that leverages large language models (LLMs) as spatio-temporal predictors and trajectory reasoners. RHYTHM partitions trajectories into…

计算与语言 · 计算机科学 2025-10-01 Haoyu He , Haozheng Luo , Yan Chen , Qi R. Wang

Motivated by recent problems in mathematical cosmology, in which temporal averaging methods are applied in order to analyze the future asymptotics of models which exhibit oscillatory behavior, we provide a theorem concerning the large-time…

动力系统 · 数学 2021-03-03 David Fajman , Gernot Heißel , Jin Woo Jang

This paper proposes a speech rhythm-based method for speaker embeddings to model phoneme duration using a few utterances by the target speaker. Speech rhythm is one of the essential factors among speaker characteristics, along with acoustic…

声音 · 计算机科学 2024-02-13 Kenichi Fujita , Atsushi Ando , Yusuke Ijima

This paper presents methods for building speech recognizers tailored for Japanese speaking assessment tasks. Specifically, we build a speech recognizer that outputs phonemic labels with accent markers. Although Japanese is resource-rich,…

计算与语言 · 计算机科学 2025-09-26 Yotaro Kubo , Richard Sproat , Chihiro Taguchi , Llion Jones

Language change is a cultural evolutionary process in which variants of linguistic variables change in frequency through processes analogous to mutation, selection and genetic drift. In this work, we apply a recently-introduced method to…

计算与语言 · 计算机科学 2023-08-22 Juan Guerrero Montero , Andres Karjus , Kenny Smith , Richard A. Blythe

This study explores the temporal dynamics of language processing by examining the alignment between word representations from a pre-trained transformer-based language model, and EEG data. Using a Temporal Response Function (TRF) model, we…

计算与语言 · 计算机科学 2024-08-01 Davide Turco , Conor Houghton

We consider the problem of mining signal temporal logical requirements from a dataset of regular (good) and anomalous (bad) trajectories of a dynamical system. We assume the training set to be labeled by human experts and that we have…

人工智能 · 计算机科学 2018-08-02 Laura Nenzi , Simone Silvetti , Ezio Bartocci , Luca Bortolussi

A growing literature on speech interruptions describes how people interrupt one another with speech, but these behaviours have not yet been implemented in the design of artificial agents which interrupt. Perceptions of a prototype proactive…

人机交互 · 计算机科学 2024-05-15 Justin Edwards , Philip R. Doyle , Holly P. Branigan , Benjamin R. Cowan

Is it possible to develop a `physics of language' which can explain the spatial, temporal and social patterns we see, and which can predict future change like we forecast the weather? Such a theory is likely to involve ideas from…

物理与社会 · 物理学 2025-12-22 James Burridge

The transduction process that occurs in the inner ear of the auditory system is a complex mechanism which requires a non-linear dynamical description. In addition to this, the stochastic phenomena that naturally arise in the inner ear…

物理教育 · 物理学 2022-01-17 Francesco Veronesi , Edoardo Milotti

Kuramoto's differential equation describes a synchronization process between several harmonic oscillators. It has been used to model biological phenomena such as the synchronization of heart cells, the circadian rhythm, or brain waves. It…

动力系统 · 数学 2026-05-26 Daniel Burns , Gregorio Malajovich , Charles Pugh , Indika Rajapakse , Steve Smale

We propose a model of the speech perception of individual words in the presence of mishearings. This phenomenological approach is based on concepts used in linguistics, and provides a formalism that is universal across languages. We put…

计算与语言 · 计算机科学 2020-10-19 Anita Mehta , Jean-Marc Luck

A method is presented for the rhythmic parsing problem: Given a sequence of observed musical note onset times, we estimate the corresponding notated rhythm and tempo process. A graphical model is developed that represents the simultaneous…

人工智能 · 计算机科学 2013-01-14 Christopher S Raphael

Static word embeddings that represent words by a single vector cannot capture the variability of word meaning in different linguistic and extralinguistic contexts. Building on prior work on contextualized and dynamic word embeddings, we…

计算与语言 · 计算机科学 2021-06-09 Valentin Hofmann , Janet B. Pierrehumbert , Hinrich Schütze

The study of motion in animals and robots has been aided by insights from geometric mechanics. In friction dominated systems, a mechanical "connection" can provide a high fidelity mechanical model. The connection is a co-vector (Lie…

生物物理 · 物理学 2018-01-26 Brian A. Bittner , Ross L. Hatton , Shai Revzen

The identification and modeling of time-varying systems is a fundamental challenge in signal processing and system identification. To address this challenge, we propose a class of time-varying state-space model (SSM) based neural networks…

机器学习 · 计算机科学 2026-05-18 Sanja Karilanova , Subhrakanti Dey , Ayça Özçelikkale
‹ 上一页 1 8 9 10 下一页 ›