中文
相关论文

相关论文: Significance of Chirp MFCC as a Feature in Speech …

200 篇论文

Analysis of signals with oscillatory modes with crossover instantaneous frequencies is a challenging problem in time series analysis. One way to handle this problem is lifting the 2-dimensional time-frequency representation to a…

数值分析 · 数学 2022-06-22 Ziyu Chen , Hau-Tieng Wu

Spectrogram-based representations have grown to dominate the feature space for deep learning audio analysis systems, and are often adopted for speech analysis also. Initially, the primary motivator for spectrogram-based representations was…

音频与语音处理 · 电气工程与系统科学 2026-03-17 Ian McLoughlin , Lam Pham , Yan Song , Xiaoxiao Miao , Huy Phan , Pengfei Cai , Qing Gu , Jiang Nan , Haoyu Song , Donny Soh

Anti-spoofing is the task of speech authentication. That is, identifying genuine human speech compared to spoofed speech. The main focus of this paper is to suggest new representations for genuine and spoofed speech, based on the…

音频与语音处理 · 电气工程与系统科学 2022-10-28 Matan Karo , Arie Yeredor , Itshak Lapidot

Animal vocalisations contain important information about health, emotional state, and behaviour, thus can be potentially used for animal welfare monitoring. Motivated by the spectro-temporal patterns of chick calls in the time$-$frequency…

音频与语音处理 · 电气工程与系统科学 2021-10-11 Changhong Wang , Emmanouil Benetos , Shuge Wang , Elisabetta Versace

To meet the increasingly demanding quality-of-service requirements of the next-generation multi-carrier mobile networks, it is essential to design multi-functional signalling schemes facilitating efficient, flexible, and reliable…

信号处理 · 电气工程与系统科学 2025-08-11 Zeping Sui , Qu Luo , Zilong Liu , Murat Temiz , Leila Musavian , Christos Masouros , Yong Liang Guan , Pei Xiao , Lajos Hanzo

In this paper, we use several techniques with conventional vocal feature extraction (MFCC, STFT), along with deep-learning approaches such as CNN, and also context-level analysis, by providing the textual data, and combining different…

音频与语音处理 · 电气工程与系统科学 2019-05-22 Andrew Huang , Puwei Bao

Recent advances in remote heart rate measurement, motivated by data-driven approaches, have notably enhanced accuracy. However, these improvements primarily focus on recovering the rPPG signal, overlooking the implicit challenges of…

计算机视觉与模式识别 · 计算机科学 2024-03-12 Joaquim Comas , Adria Ruiz , Federico Sukno

Recent successful applications of convolutional neural networks (CNNs) to audio classification and speech recognition have motivated the search for better input representations for more efficient training. Visual displays of an audio…

计算机视觉与模式识别 · 计算机科学 2017-06-23 M. Huzaifah

Speaker verification systems have seen significant advancements with the introduction of Multi-scale Feature Aggregation (MFA) architectures, such as MFA-Conformer and ECAPA-TDNN. These models leverage information from various network…

声音 · 计算机科学 2024-10-08 Satvik Dixit , Massa Baali , Rita Singh , Bhiksha Raj

We propose a multiple chirp rate index modulation (MCR-IM) system based on Zadoff-Chu (ZC) sequences that overcomes the problems of low transmission rate and large-scale access in classical LoRa networks. We demonstrate the extremely low…

信号处理 · 电气工程与系统科学 2025-07-22 Xiaobin Zhu , Minling Zhang , Guofa Cai , Jiguang He , Georges Kaddoum

Speech recognition and speaker identification are important for authentication and verification in security purpose, but they are difficult to achieve. Speaker identification methods can be divided into text-independent and text-dependent.…

机器学习 · 计算机科学 2010-09-28 S. M. Kamruzzaman , A. N. M. Rezaul Karim , Md. Saiful Islam , Md. Emdadul Haque

We propose the multi-layered cepstrum (MLC) method to estimate multiple fundamental frequencies (MF0) of a signal under challenging contamination such as high-pass filter noise. Taking the operation of cepstrum (i.e., Fourier transform,…

音频与语音处理 · 电气工程与系统科学 2019-02-05 Chin-Yun Yu , Li Su

A new musical instrument classification method using convolutional neural networks (CNNs) is presented in this paper. Unlike the traditional methods, we investigated a scheme for classifying musical instruments using the learned features…

声音 · 计算机科学 2015-12-24 Taejin Park , Taejin Lee

This paper proposes a method for generating speech from filterbank mel frequency cepstral coefficients (MFCC), which are widely used in speech applications, such as ASR, but are generally considered unusable for speech synthesis. First, we…

音频与语音处理 · 电气工程与系统科学 2018-04-04 Lauri Juvela , Bajibabu Bollepalli , Xin Wang , Hirokazu Kameoka , Manu Airaksinen , Junichi Yamagishi , Paavo Alku

This paper presents a fully automated approach for identifying speech anomalies from voice recordings to aid in the assessment of speech impairments. By combining Connectionist Temporal Classification (CTC) and encoder-decoder-based…

声音 · 计算机科学 2023-08-04 Laurin Wagner , Mario Zusag , Theresa Bloder

Emotion recognition from audio signals has been regarded as a challenging task in signal processing as it can be considered as a collection of static and dynamic classification tasks. Recognition of emotions from speech data has been…

声音 · 计算机科学 2020-09-21 Soham Chattopadhyay , Arijit Dey , Hritam Basak

This paper introduces scattering transform for speech emotion recognition (SER). Scattering transform generates feature representations which remain stable to deformations and shifting in time and frequency without much loss of information.…

音频与语音处理 · 电气工程与系统科学 2021-05-12 Premjeet Singh , Goutam Saha , Md Sahidullah

The previous SpEx+ has yielded outstanding performance in speaker extraction and attracted much attention. However, it still encounters inadequate utilization of multi-scale information and speaker embedding. To this end, this paper…

声音 · 计算机科学 2023-06-29 Jun Chen , Wei Rao , Zilin Wang , Jiuxin Lin , Yukai Ju , Shulin He , Yannan Wang , Zhiyong Wu

Homomorphic analysis is a well-known method for the separation of non-linearly combined signals. More particularly, the use of complex cepstrum for source-tract deconvolution has been discussed in various articles. However there exists no…

声音 · 计算机科学 2020-01-01 Thomas Drugman , Baris Bozkurt , Thierry Dutoit

We present a novel multicarrier waveform, termed chirp-permuted affine frequency division multiplexing (CP-AFDM), which introduces a unique chirp-permutation domain on top of the chirp subcarriers of the conventional AFDM. Rigorous analysis…

信号处理 · 电气工程与系统科学 2025-09-24 Hyeon Seok Rou , Giuseppe Thadeu Freitas de Abreu