English
Related papers

Related papers: Spoken Digit Recognition and Speaker Classificatio…

200 papers

As a test of general applicability, we use the recently proposed spin-wave delay line active-ring reservoir computer to perform the spoken digit recognition task. On this, classification accuracies of up to 93% are achieved. The tested…

Neural and Evolutionary Computing · Computer Science 2020-05-27 Stuart Watt , Mikhail Kostylev

Physical reservoir computing, which is a promising method for the implementation of highly efficient artificial intelligence devices, requires a physical system with nonlinearity, fading memory, and the ability to map in high dimensions.…

Emerging Technologies · Computer Science 2022-07-08 Wataru Namiki , Daiki Nishioka , Yu Yamaguchi , Takashi Tsuchiya , Tohru Higuchi , Kazuya Terabe

Physical reservoir computing (PRC) is a promising brain-inspired computing architecture for overcoming the von Neumann bottleneck by utilizing the intrinsic dynamics of physical systems. However, a major obstacle to its real-world…

Emerging Technologies · Computer Science 2026-03-06 Jiaxuan Chen , Ryo Iguchi , Sota Hikasa , Takashi Tsuchiya

The reservoir computing neural network architecture is widely used to test hardware systems for neuromorphic computing. One of the preferred tasks for bench-marking such devices is automatic speech recognition. However, this task requires…

Speech recognition is a critical task in the field of artificial intelligence and has witnessed remarkable advancements thanks to large and complex neural networks, whose training process typically requires massive amounts of labeled data…

Neural and Evolutionary Computing · Computer Science 2024-05-24 Enrico Picco , Alessandro Lupo , Serge Massar

Reservoir computing is a brain-inspired machine learning framework for processing temporal data by mapping inputs into high-dimensional spaces. Physical reservoir computers (PRCs) leverage native fading memory and nonlinearity in physical…

Emerging Technologies · Computer Science 2024-05-16 Ahmed S. Mohamed , Anurag Dhungel , Md Sakib Hasan , Joseph S. Najem

Physical Reservoir Computing (PRC) is an unconventional computing paradigm, which exploits nonlinear dynamics of reservoir blocks to perform recognition and classification tasks. Here we show with simulations that patterned thin films…

Mesoscale and Nanoscale Physics · Physics 2023-05-18 Md Mahadi Rajib , Walid Al Misba , Md. Fahim F. Chowdhury , Muhammad Sabbir Alam , Jayasimha Atulasimha

Alzheimer's disease and related dementias (ADRD) affect one in five adults over 60, yet more than half of individuals with cognitive decline remain undiagnosed. Speech-based assessments show promise for early detection, as phonetic motor…

Speaker recognition is a biometric modality that uses underlying speech information to determine the identity of the speaker. Speaker Identification (SID) under noisy conditions is one of the challenging topics in the field of speech…

Sound · Computer Science 2019-08-02 Nursadul Mamun , Ria Ghosh , John H. L. Hansen

In this work, we explore the possibility of decoding Imagined Speech brain waves using machine learning techniques. We propose a covariance matrix of Electroencephalogram channels as input features, projection to tangent space of covariance…

Signal Processing · Electrical Eng. & Systems 2021-05-03 Abhiram Singh , Ashwin Gumaste

While physical reservoir computing (PRC) is a promising way to achieve low power consumption neuromorphic computing, its computational performance is still insufficient at a practical level. One promising approach to improving PRC…

Applied Physics · Physics 2023-09-12 Daiki Nishioka , Takashi Tsuchiya , Masataka Imura , Yasuo Koide , Tohru Higuchi , Kazuya Terabe

Automatic speech recognition (ASR) plays a pivotal role in our daily lives, offering utility not only for interacting with machines but also for facilitating communication for individuals with partial or profound hearing impairments. The…

Audio and Speech Processing · Electrical Eng. & Systems 2025-01-14 Billel Essaid , Hamza Kheddar , Noureddine Batel , Muhammad E. H. Chowdhury , Abderrahmane Lakas

In this paper we extend our earlier work of (Rietman et al. 2022) presenting an application of physical Reservoir Computing (RC) to the classification of handwritten and spoken digits. We utilize an unpoled cube of Lead Zirconate Titanate…

Machine Learning · Computer Science 2026-04-02 Thomas Buckley , Leslie Schumm , Manor Askenazi , Edward Rietman

Despite the advancements in cutting-edge technologies, audio signal processing continues to pose challenges and lacks the precision of a human speech processing system. To address these challenges, we propose a novel approach to simplify…

Sound · Computer Science 2026-03-26 Rinku Sebastian , Simon O'Keefe , Martin Trefzer

This study examines the effectiveness of traditional machine learning classifiers versus deep learning models for detecting the imagined speech using electroencephalogram data. Specifically, we evaluated conventional machine learning…

Machine Learning · Computer Science 2024-12-18 Byung-Kwan Ko , Jun-Young Kim , Seo-Hyun Lee

The intrinsic dynamics and event-driven nature of spiking neural networks (SNNs) make them excel in processing temporal information by naturally utilizing embedded time sequences as time steps. Recent studies adopting this approach have…

Machine Learning · Computer Science 2024-12-18 Jiaqi Wang , Liutao Yu , Liwei Huang , Chenlin Zhou , Han Zhang , Zhenxi Song , Min Zhang , Zhengyu Ma , Zhiguo Zhang

This paper deals the combination of nonlinear predictive models with classical LPCC parameterization for speaker recognition. It is shown that the combination of both a measure defined over LPCC coefficients and a measure defined over…

Sound · Computer Science 2022-03-08 Marcos Faundez-Zanuy

Gradient clipping plays a vital role in training large-scale automatic speech recognition (ASR) models. It is typically applied to minibatch gradients to prevent gradient explosion, and to the individual sample gradients to mitigate…

Cryptography and Security · Computer Science 2024-06-07 Lun Wang , Om Thakkar , Zhong Meng , Nicole Rafidi , Rohit Prabhavalkar , Arun Narayanan

Over the last decade, deep-learning methods have been gradually incorporated into conventional automatic speech recognition (ASR) frameworks to create acoustic, pronunciation, and language models. Although it led to significant improvements…

Sound · Computer Science 2022-05-26 Zohreh Ansari , Farzin Pourhoseini , Fatemeh Hadaeghi

Physical reservoir computing is a computational paradigm that enables spatio-temporal pattern recognition to be performed directly in matter. The use of physical matter leads the way towards energy-efficient devices capable of solving…

Mesoscale and Nanoscale Physics · Physics 2025-07-08 Robin Msiska , Jake Love , Jeroen Mulkers , Jonathan Leliaert , Karin Everschor-Sitte
‹ Prev 1 2 3 10 Next ›