相关论文: Model of the Songbird Nucleus HVC as a Network of …
The means by which neuronal activity yields robust behavior is a ubiquitous question in neuroscience. In the songbird, the timing of a highly stereotyped song motif is attributed to the cortical nucleus HVC, and to feedback to HVC from…
Complex, learned motor behaviors involve the coordination of large-scale neural activity across multiple brain regions, but our understanding of the population-level dynamics within different regions tied to the same behavior remains…
We demonstrate numerically that a brief burst consisting of two to six spikes can propagate in a stable manner through a one-dimensional homogeneous feedforward chain of non-bursting neurons with excitatory synaptic connections. Our results…
Behavioral sequences of animals are often structured and can be described by probabilistic rules (or "action syntax"). The patterns of vocal elements in birdsong are a prime example. The encoding of such rules in neural circuits is poorly…
With the goal of building a model of the HVC nucleus in the avian song system, we discuss in detail a model of HVC$_{\text{RA}}$ projection neurons comprised of a somatic compartment with fast Na$^+$ and K$^+$ currents and a dendritic…
We propose another integrate-and-fire model as a single neuron model. We study a globally coupled noisy integrate-and-fire model with inhibitory interaction using the Fokker-Planck equation and the Langevin equation, and find a reentrant…
Typically, singing voice conversion (SVC) depends on an embedding vector, extracted from either a speaker lookup table (LUT) or a speaker recognition network (SRN), to model speaker identity. However, singing contains more expressive…
It is challenging to accelerate the training process while ensuring both high-quality generated voices and acceptable inference speed. In this paper, we propose a novel neural vocoder called InstructSing, which can converge much faster…
The problem of deciphering how low-level patterns (action potentials in the brain, amino acids in a protein, etc.) drive high-level biological features (sensorimotor behavior, enzymatic function) represents the central challenge of…
We study the stability and information encoding capacity of synchronized states in a neuronal network model that represents part of thalamic circuitry. Our model neurons have a Hodgkin-Huxley-type low threshold Calcium channel, display post…
Practice of a complex motor gesture involves exploration of motor space to attain a better match to target output, but little is known about the neural code for such exploration. Here, we examine spiking in an area of the songbird brain…
Singing Voice Conversion (SVC) has emerged as a significant subfield of Voice Conversion (VC), enabling the transformation of one singer's voice into another while preserving musical elements such as melody, rhythm, and timbre. Traditional…
In the classic view of cortical rhythms, the interaction between excitatory pyramidal neurons (E) and inhibitory parvalbumin neurons (I) has been shown to be sufficient to generate gamma and beta band rhythms. However, it is now clear that…
This paper proposes a data-efficient, semi-supervised, two-pass framework for segmenting bird vocalizations. The framework utilizes a binary classification model to categorize frames of an input audio recording into the background or bird…
We investigate the dynamical role of inhibitory and highly connected nodes (hub) in synchronization and input processing of leaky-integrate-and-fire neural networks with short term synaptic plasticity. We take advantage of a heterogeneous…
Singing voice conversion (SVC) is hindered by noise sensitivity due to the use of non-robust methods for extracting pitch and energy during the inference. As clean signals are key for the source audio in SVC, music source separation…
Consecutive repetition of actions is common in behavioral sequences. Although integration of sensory feedback with internal motor programs is important for sequence generation, if and how feedback contributes to repetitive actions is poorly…
This paper presents FastSVC, a light-weight cross-domain singing voice conversion (SVC) system, which can achieve high conversion performance, with inference speed 4x faster than real-time on CPUs. FastSVC uses Conformer-based phoneme…
Singing voice conversion (SVC) is one promising technique which can enrich the way of human-computer interaction by endowing a computer the ability to produce high-fidelity and expressive singing voice. In this paper, we propose DiffSVC, an…
Singing Voice Conversion (SVC) transfers a source singer's timbre to a target while keeping melody and lyrics. The key challenge in any-to-any SVC is adapting unseen speaker timbres to source audio without quality degradation. Existing…