中文
相关论文

相关论文: Vowel recognition with four coupled spin-torque na…

200 篇论文

Present day computers expend orders of magnitude more computational resources to perform various cognitive and perception related tasks that humans routinely perform everyday. This has recently resulted in a seismic shift in the field of…

新兴技术 · 计算机科学 2017-12-22 Abhronil Sengupta , Kaushik Roy

In this paper, we propose an innovative approach to perform speaker recognition by fusing two recently introduced deep neural networks (DNNs) namely - SincNet and X-Vector. The idea behind using SincNet filters on the raw speech waveform is…

计算与语言 · 计算机科学 2020-04-07 Mayank Tripathi , Divyanshu Singh , Seba Susan

Motivated by the aim to find new medical strategies to suppress undesirable neural synchronization we study the control of oscillations in a system of inhibitory coupled noisy oscillators. Using dynamical properties of inhibition, we find…

无序系统与神经网络 · 物理学 2009-05-27 C. J. Tessone , E. Ullner , A. A. Zaikin , J. Kurths , R. Toral

Neuromorphic computing is an emerging technology enabling low-latency and energy-efficient signal processing. A key algorithmic tool in neuromorphic computing is spiking neural networks (SNNs). SNNs are biologically inspired neural networks…

机器学习 · 计算机科学 2025-08-11 Sanja Karilanova , Subhrakanti Dey , Ayça Özçelikkale

We investigate the synchronization of oscillators based on anharmonic nanoelectromechanical resonators. Our experimental implementation allows unprecedented observation and control of parameters governing the dynamics of synchronization. We…

介观与纳米尺度物理 · 物理学 2014-01-15 M. H. Matheny , M. Grau , L. G. Villanueva , R. B. Karabalin , M. C. Cross , M. L. Roukes

Recent advances in unsupervised speech representation learning discover new approaches and provide new state-of-the-art for diverse types of speech processing tasks. This paper presents an investigation of using wav2vec 2.0 deep speech…

The oscillatory response of nonlinear systems exhibits characteristic phenomena such as multistability, discontinuous jumps and hysteresis. These can be utilized in applications leading, e.g., to precise frequency measurement, mixing,…

介观与纳米尺度物理 · 物理学 2015-05-14 Quirin P. Unterreithmeier , Thomas Faust , Jorg P. Kotthaus

The dynamics of vortex based spin-torque nano-oscillators is investigated theoretically. Starting from a fully analytical model based on the Thiele equation approach, fine-tuned data-driven corrections are carried out to the gyrotropic and…

介观与纳米尺度物理 · 物理学 2022-06-29 Flavio Abreu Araujo , Chloé Chopin , Simon De Wergifosse

We present a transformer-based architecture for voice separation of a target speaker from multiple other speakers and ambient noise. We achieve this by using two separate neural networks: (A) An enrolment network designed to craft…

音频与语音处理 · 电气工程与系统科学 2025-01-03 Akam Rahimi , Triantafyllos Afouras , Andrew Zisserman

Pre-trained vision models have found widespread application across diverse domains. Prompt tuning-based methods have emerged as a parameter-efficient paradigm for adapting pre-trained vision models. While effective on standard benchmarks,…

计算机视觉与模式识别 · 计算机科学 2026-04-21 Qiugang Zhan , Anning Jiang , Ran Tao , Ao Ma , Xiangyu Zhang , Xiurui Xie , Guisong Liu

Understanding how the brain learns to compute functions reliably, efficiently and robustly with noisy spiking activity is a fundamental challenge in neuroscience. Most sensory and motor tasks can be described as dynamical systems and could…

神经元与认知 · 定量生物学 2017-05-24 Sophie Denève , Alireza Alemi , Ralph Bourdoukan

Temporal coding is one approach to representing information in spiking neural networks. An example of its application is the location of sounds by barn owls that requires especially precise temporal coding. Dependent upon the azimuthal…

神经元与认知 · 定量生物学 2014-01-24 Thomas Pfeil , Anne-Christine Scherzer , Johannes Schemmel , Karlheinz Meier

Speech applications are expected to be low-power and robust under noisy conditions. An effective Voice Activity Detection (VAD) front-end lowers the computational need. Spiking Neural Networks (SNNs) are known to be biologically plausible…

声音 · 计算机科学 2024-03-12 Qu Yang , Qianhui Liu , Nan Li , Meng Ge , Zeyang Song , Haizhou Li

Action potentials are the basic unit of information in the nervous system and their reliable detection and decoding holds the key to understanding how the brain generates complex thought and behavior. Transducing these signals into…

Current speech production systems predominantly rely on large transformer models that operate as black boxes, providing little interpretability or grounding in the physical mechanisms of human speech. We address this limitation by proposing…

音频与语音处理 · 电气工程与系统科学 2025-10-08 Akshay Anand , Chenxu Guo , Cheol Jun Cho , Jiachen Lian , Gopala Anumanchipalli

Despite the remarkable progress in the synthesis speed and fidelity of neural vocoders, their high energy consumption remains a critical barrier to practical deployment on computationally restricted edge devices. Spiking Neural Networks…

机器学习 · 计算机科学 2025-09-17 Yukun Chen , Zhaoxi Mu , Andong Li , Peilin Li , Xinyu Yang

A spintronic method of ultra-fast broadband microwave spectrum analysis is proposed. It uses a rapidly tuned spin torque nano-oscillator (STNO), and does not require injection locking. This method treats an STNO generating a microwave…

Spiking neural networks (SNNs) are the third generation of neural networks that are biologically inspired to process data in a fashion that emulates the exchange of signals in the brain. Within the Computer Vision community SNNs have…

音频与语音处理 · 电气工程与系统科学 2024-09-04 William Bjorndahl , Jack Easton , Austin Modoff , Eric C. Larson , Joseph Camp , Prasanna Rangarajan

We study experimentally the dynamic tunability and self-induced nonlinearity of split-ring resonators incorporating variable capacitance diodes. We demonstrate that the eigenfrequencies of the resonators can be tuned over a wide frequency…

光学 · 物理学 2009-11-13 Ilya V. Shadrivov , Steven K. Morrison , Yuri S. Kivshar

While Word2Vec represents words (in text) as vectors carrying semantic information, audio Word2Vec was shown to be able to represent signal segments of spoken words as vectors carrying phonetic structure information. Audio Word2Vec can be…

计算与语言 · 计算机科学 2018-08-08 Yu-Hsuan Wang , Hung-yi Lee , Lin-shan Lee