English

Meta-Learning in Audio and Speech Processing: An End to End Comprehensive Review

Sound 2025-03-14 v1 Machine Learning Audio and Speech Processing

Abstract

This survey overviews various meta-learning approaches used in audio and speech processing scenarios. Meta-learning is used where model performance needs to be maximized with minimum annotated samples, making it suitable for low-sample audio processing. Although the field has made some significant contributions, audio meta-learning still lacks the presence of comprehensive survey papers. We present a systematic review of meta-learning methodologies in audio processing. This includes audio-specific discussions on data augmentation, feature extraction, preprocessing techniques, meta-learners, task selection strategies and also presents important datasets in audio, together with crucial real-world use cases. Through this extensive review, we aim to provide valuable insights and identify future research directions in the intersection of meta-learning and audio processing.

Keywords

Cite

@article{arxiv.2408.10330,
  title  = {Meta-Learning in Audio and Speech Processing: An End to End Comprehensive Review},
  author = {Athul Raimon and Shubha Masti and Shyam K Sateesh and Siyani Vengatagiri and Bhaskarjyoti Das},
  journal= {arXiv preprint arXiv:2408.10330},
  year   = {2025}
}

Comments

Survey Paper (15 pages, 1 figure)

R2 v1 2026-06-28T18:17:20.405Z