中文
相关论文

相关论文: Speech-based Age and Gender Prediction with Transf…

200 篇论文

This paper introduces a general classifier based on WavLM features, to infer demographic characteristics, such as age, gender, native language, education, and country, from speech. Demographic feature prediction plays a crucial role in…

计算与语言 · 计算机科学 2025-02-18 Yuchen Yang , Thomas Thebaud , Najim Dehak

The estimation of speaker characteristics such as age and height is a challenging task, having numerous applications in voice forensic analysis. In this work, we propose a bi-encoder transformer mixture model for speaker age and height…

声音 · 计算机科学 2022-03-23 Tarun Gupta , Duc-Tuan Truong , Tran The Anh , Chng Eng Siong

Children have less text understanding capability than adults. Moreover, this capability differs among the children of different ages. Hence, automatically predicting a recommended age based on texts or sentences would be a great benefit to…

计算与语言 · 计算机科学 2023-08-22 Rashedur Rahman , Gwénolé Lecorvé , Nicolas Béchet

In this paper we extend the x-vector framework for the task of speaker's age estimation and gender classification. In particular, we replace the baseline multilayer-TDNN architecture with QuartzNet, a convolutional architecture that has…

音频与语音处理 · 电气工程与系统科学 2020-12-04 Damian Kwasny , Daria Hemmerling

This paper is focused on the finetuning of acoustic models for speaker adaptation goals on a given gender. We pretrained the Transformer baseline model on Librispeech-960 and conduct experiments with finetuning on the gender-specific test…

音频与语音处理 · 电气工程与系统科学 2020-11-18 Sokolov Artem , Andrey V. Savchenko

The interpretation of human voices holds importance across various applications. This study ventures into predicting age, gender, and emotion from vocal cues, a field with vast applications. Voice analysis tech advancements span domains,…

音频与语音处理 · 电气工程与系统科学 2024-03-05 Aron R , Indra Sigicharla , Chirag Periwal , Mohanaprasad K , Nithya Darisini P S , Sourabh Tiwari , Shivani Arora

VoxCeleb datasets are widely used in speaker recognition studies. Our work serves two purposes. First, we provide speaker age labels and (an alternative) annotation of speaker gender. Second, we demonstrate the use of this metadata by…

机器学习 · 计算机科学 2021-12-21 Khaled Hechmi , Trung Ngo Trong , Ville Hautamaki , Tomi Kinnunen

In this project, competition-winning deep neural networks with pretrained weights are used for image-based gender recognition and age estimation. Transfer learning is explored using both VGG19 and VGGFace pretrained models by testing the…

计算机视觉与模式识别 · 计算机科学 2018-11-20 Philip Smith , Cuixian Chen

Children's speech presents challenges for age and gender classification due to high variability in pitch, articulation, and developmental traits. While self-supervised learning (SSL) models perform well on adult speech tasks, their ability…

音频与语音处理 · 电气工程与系统科学 2025-08-15 Abhijit Sinha , Harishankar Kumar , Mohit Joshi , Hemant Kumar Kathania , Shrikanth Narayanan , Sudarsana Reddy Kadiri

While unsupervised variational autoencoders (VAE) have become a powerful tool in neuroimage analysis, their application to supervised learning is under-explored. We aim to close this gap by proposing a unified probabilistic model for…

机器学习 · 计算机科学 2019-07-15 Qingyu Zhao , Ehsan Adeli , Nicolas Honnorat , Tuo Leng , Kilian M. Pohl

The human brain's white matter (WM) structure is of immense interest to the scientific community. Diffusion MRI gives a powerful tool to describe the brain WM structure noninvasively. To potentially enable monitoring of age-related changes…

图像与视频处理 · 电气工程与系统科学 2022-02-09 Hao He , Fan Zhang , Steve Pieper , Nikos Makris , Yogesh Rathi , William Wells , Lauren J. O'Donnell

In this paper, we demonstrated the benefit of using pre-trained model to extract acoustic embedding to jointly predict (multitask learning) three tasks: emotion, age, and native country. The pre-trained model was trained with wav2vec 2.0…

音频与语音处理 · 电气工程与系统科学 2022-09-28 Bagus Tris Atmaja , Zanjabila , Akira Sasou

Realistic age-progressed photos provide invaluable biometric information in a wide range of applications. In recent years, deep learning-based approaches have made remarkable progress in modeling the aging process of the human face.…

计算机视觉与模式识别 · 计算机科学 2020-11-17 Yao Xiao , Yijun Zhao

Objectives: Age and gender estimation is crucial for various applications, including forensic investigations and anthropological studies. This research aims to develop a predictive system for age and gender estimation in living individuals,…

Self-supervised models for speech processing emerged recently as popular foundation blocks in speech processing pipelines. These models are pre-trained on unlabeled audio data and then used in speech processing downstream tasks such as…

计算与语言 · 计算机科学 2022-07-06 Marcely Zanon Boito , Laurent Besacier , Natalia Tomashenko , Yannick Estève

Automatic prediction of age and gender from face images has drawn a lot of attention recently, due it is wide applications in various facial analysis problems. However, due to the large intra-class variation of face images (such as…

计算机视觉与模式识别 · 计算机科学 2020-12-09 Amirali Abdolrashidi , Mehdi Minaei , Elham Azimi , Shervin Minaee

Purpose: To develop an age prediction model which is interpretable and robust to demographic and technological variances in brain MRI scans. Materials and Methods: We propose a transformer-based architecture that leverages self-supervised…

计算机视觉与模式识别 · 计算机科学 2025-06-24 Pengyu Kan , Craig Jones , Kenichi Oishi

Human age estimation has attracted increasing researches due to its wide applicability in such as security monitoring and advertisement recommendation. Although a variety of methods have been proposed, most of them focus only on the…

计算机视觉与模式识别 · 计算机科学 2016-09-14 Qing Tian , Songcan Chen , Xiaoyang Tan

Multilingual speech recognition with supervised learning has achieved great results as reflected in recent research. With the development of pretraining methods on audio and text data, it is imperative to transfer the knowledge from…

计算与语言 · 计算机科学 2022-05-26 Ngoc-Quan Pham , Alex Waibel , Jan Niehues

Recently proposed self-supervised learning approaches have been successful for pre-training speech representation models. The utility of these learned representations has been observed empirically, but not much has been studied about the…

计算与语言 · 计算机科学 2022-12-06 Ankita Pasad , Ju-Chieh Chou , Karen Livescu
‹ 上一页 1 2 3 10 下一页 ›