English

Generating Segment Durations in a Text-To-Speech System: A Hybrid Rule-Based/Neural Network Approach

Neural and Evolutionary Computing 2007-05-23 v1 Human-Computer Interaction

Abstract

A combination of a neural network with rule firing information from a rule-based system is used to generate segment durations for a text-to-speech system. The system shows a slight improvement in performance over a neural network system without the rule firing information. Synthesized speech using segment durations was accepted by listeners as having about the same quality as speech generated using segment durations extracted from natural speech.

Keywords

Cite

@article{arxiv.cs/9811030,
  title  = {Generating Segment Durations in a Text-To-Speech System: A Hybrid Rule-Based/Neural Network Approach},
  author = {Gerald Corrigan and Noel Massey and Orhan Karaali},
  journal= {arXiv preprint arXiv:cs/9811030},
  year   = {2007}
}

Comments

4 pages, PostScript

R2 v1 2026-07-22T12:28:54.226Z