English

HiFi-Stream: Streaming Speech Enhancement with Generative Adversarial Networks

Sound 2025-07-24 v2 Machine Learning Audio and Speech Processing

Abstract

Speech Enhancement techniques have become core technologies in mobile devices and voice software. Still, modern deep learning solutions often require high amount of computational resources what makes their usage on low-resource devices challenging. We present HiFi-Stream, an optimized version of recently published HiFi++ model. Our experiments demonstrate that HiFi-Stream saves most of the qualities of the original model despite its size and computational complexity improved in comparison to the original HiFi++ making it one of the smallest and fastest models available. The model is evaluated in streaming setting where it demonstrates its superior performance in comparison to modern baselines.

Keywords

Cite

@article{arxiv.2503.17141,
  title  = {HiFi-Stream: Streaming Speech Enhancement with Generative Adversarial Networks},
  author = {Ekaterina Dmitrieva and Maksim Kaledin},
  journal= {arXiv preprint arXiv:2503.17141},
  year   = {2025}
}

Comments

5 pages (4 content pages + 1 page of references)

R2 v1 2026-06-28T22:29:44.022Z