中文
相关论文

相关论文: CL-UZH submission to the NIST SRE 2024 Speaker Rec…

200 篇论文

The NIST Speaker Recognition Evaluation - Conversational Telephone Speech (CTS) challenge 2019 was an open evaluation for the task of speaker verification in challenging conditions. In this paper, we provide a detailed account of the LEAP…

音频与语音处理 · 电气工程与系统科学 2020-05-26 Shreyas Ramoji , Prashant Krishnan , Bhargavram Mysore , Prachi Singh , Sriram Ganapathy

In this report, we describe the submission of Brno University of Technology (BUT) team to the VoxCeleb Speaker Recognition Challenge (VoxSRC) 2019. We also provide a brief analysis of different systems on VoxCeleb-1 test sets. Submitted…

音频与语音处理 · 电气工程与系统科学 2019-10-29 Hossein Zeinali , Shuai Wang , Anna Silnova , Pavel Matějka , Oldřich Plchot

This report describes our speaker verification systems for the tasks of the CN-Celeb Speaker Recognition Challenge 2022 (CNSRC 2022). This challenge includes two tasks, namely speaker verification(SV) and speaker retrieval(SR). The SV task…

声音 · 计算机科学 2022-09-23 Yu Zheng , Yihao Chen , Jinghan Peng , Yajun Zhang , Min Liu , Minqiang Xu

In this report, we describe our submitted system for track 2 of the VoxCeleb Speaker Recognition Challenge 2022 (VoxSRC-22). We fuse a variety of good-performing models ranging from supervised models to self-supervised learning(SSL)…

声音 · 计算机科学 2022-09-26 Gang Liu , Tianyan Zhou , Yong Zhao , Yu Wu , Zhuo Chen , Yao Qian , Jian Wu

This technical report describes the SJTU X-LANCE Lab system for the three tracks in CNSRC 2022. In this challenge, we explored the speaker embedding modeling ability of deep ResNet (Deeper r-vector). All the systems are only trained on the…

声音 · 计算机科学 2023-05-16 Zhengyang Chen , Bei Liu , Bing Han , Leying Zhang , Yanmin Qian

In this report, we describe our submission to the VoxCeleb Speaker Recognition Challenge (VoxSRC) 2020. Two approaches are adopted. One is to apply query expansion on speaker verification, which shows significant progress compared to…

声音 · 计算机科学 2020-11-06 Yu-Sen Cheng , Chun-Liang Shih , Tien-Hong Lo , Wen-Ting Tseng , Berlin Chen

We present a comprehensive analysis of the embedding extractors (frontends) developed by the ABC team for the audio track of NIST SRE 2024. We follow the two scenarios imposed by NIST: using only a provided set of telephone recordings for…

This is a description of our effort in VOiCES 2019 Speaker Recognition challenge. All systems in the fixed condition are based on the x-vector paradigm with different features and DNN topologies. The single best system reaches 1.2% EER and…

音频与语音处理 · 电气工程与系统科学 2019-07-16 Hossein Zeinali , Pavel Matějka , Ladislav Mošner , Oldřich Plchot , Anna Silnova , Ondřej Novotný , Ján Profant , Ondřej Glembek , Lukáš Burget

In this paper, we describe the top-scoring submissions for team RTZR VoxCeleb Speaker Recognition Challenge 2022 (VoxSRC-22) in the closed dataset, speaker verification Track 1. The top performed system is a fusion of 7 models, which…

音频与语音处理 · 电气工程与系统科学 2022-09-22 Sangwon Suh , Sunjong Park

This paper presents a description of STC Ltd. systems submitted to the NIST 2021 Speaker Recognition Evaluation for both fixed and open training conditions. These systems consists of a number of diverse subsystems based on using deep neural…

This paper presents the system description of the THUEE team for the NIST 2020 Speaker Recognition Evaluation (SRE) conversational telephone speech (CTS) challenge. The subsystems including ResNet74, ResNet152, and RepVGG-B2 are developed…

声音 · 计算机科学 2022-10-13 Yu Zheng , Jinghan Peng , Miao Zhao , Yufeng Ma , Min Liu , Xinyue Ma , Tianyu Liang , Tianlong Kong , Liang He , Minqiang Xu

This report describes ID R&D team submissions for Track 2 (open) to the VoxCeleb Speaker Recognition Challenge 2023 (VoxSRC-23). Our solution is based on the fusion of deep ResNets and self-supervised learning (SSL) based models trained on…

音频与语音处理 · 电气工程与系统科学 2023-08-22 Nikita Torgashov , Rostislav Makarov , Ivan Yakovlev , Pavel Malov , Andrei Balykin , Anton Okhotnikov

In this report, we describe the submission of ShaneRun's team to the VoxCeleb Speaker Recognition Challenge (VoxSRC) 2020. We use ResNet-34 as encoder to extract the speaker embeddings, which is referenced from the open-source…

声音 · 计算机科学 2020-11-04 Shen Chen

This report describes the UNISOUND submission for Track1 and Track2 of VoxCeleb Speaker Recognition Challenge 2023 (VoxSRC 2023). We submit the same system on Track 1 and Track 2, which is trained with only VoxCeleb2-dev. Large-scale ResNet…

音频与语音处理 · 电气工程与系统科学 2023-08-25 Yu Zheng , Yajun Zhang , Chuanying Niu , Yibin Zhan , Yanhua Long , Dongxing Xu

This technical report describes ChinaTelecom system for Track 1 (closed) of the VoxCeleb2023 Speaker Recognition Challenge (VoxSRC 2023). Our system consists of several ResNet variants trained only on VoxCeleb2, which were fused for better…

声音 · 计算机科学 2023-08-17 Mengjie Du , Xiang Fang , Jie Li

This report describes the systems submitted to the first and second tracks of the VoxCeleb Speaker Recognition Challenge (VoxSRC) 2020, which ranked second in both tracks. Three key points of the system pipeline are explored: (1)…

声音 · 计算机科学 2020-11-03 Xu Xiang

This paper presents the Intelligent Voice (IV) system submitted to the NIST 2016 Speaker Recognition Evaluation (SRE). The primary emphasis of SRE this year was on developing speaker recognition technology which is robust for novel…

声音 · 计算机科学 2016-11-03 Abbas Khosravani , Cornelius Glackin , Nazim Dugan , Gérard Chollet , Nigel Cannings

Many speaker recognition challenges have been held to assess the speaker verification system in the wild and probe the performance limit. Voxceleb Speaker Recognition Challenge (VoxSRC), based on the voxceleb, is the most popular. Besides,…

声音 · 计算机科学 2023-06-02 Zhengyang Chen , Bing Han , Xu Xiang , Houjun Huang , Bei Liu , Yanmin Qian

This paper delineates the visual speech recognition (VSR) system introduced by the NPU-ASLP (Team 237) in the second Chinese Continuous Visual Speech Recognition Challenge (CNVSRC 2024), engaging in all four tracks, including the fixed and…

计算机视觉与模式识别 · 计算机科学 2024-09-13 He Wang , Lei Xie

This work describes the speaker verification system developed by Human Language Technology Laboratory, National University of Singapore (HLT-NUS) for 2019 NIST Multimedia Speaker Recognition Evaluation (SRE). The multimedia research has…

音频与语音处理 · 电气工程与系统科学 2020-10-09 Rohan Kumar Das , Ruijie Tao , Jichen Yang , Wei Rao , Cheng Yu , Haizhou Li
‹ 上一页 1 2 3 10 下一页 ›