Related papers: Tunable add-drop filter using an active whispering…
There is an increasing number of pre-trained deep neural network models. However, it is still unclear how to effectively use these models for a new task. Transfer learning, which aims to transfer knowledge from source tasks to a target…
Building upon Diff-A-Riff, a latent diffusion model for musical instrument accompaniment generation, we present a series of improvements targeting quality, diversity, inference speed, and text-driven control. First, we upgrade the…
Optical energy flow inside a dielectric microsphere exposed to an optical wave is usually codirected with its wave vector. At the same time, if the optical field in a microparticle is in resonance with a high-quality spatial eigenmode,…
Whispering gallery resonators (WGR's), based on total internal reflection, possess high quality factors in a broad spectral range. Thus, nonlinear optical processes in such cavities are ideally suited for the generation of broadband or…
We have fabricated a layered nano-composite by alternating metal and gain medium layers, the gain dielectric consisting of a polymer incorporating optically pumped dye molecules. Exploiting an improved version of the effective medium…
We demonstrate a narrow line, fiber loop laser using Erbium-doped fiber as the gain material, stabilized by using a microsphere as a transmissive frequency selective element. Stable lasing with a linewidth of 170 kHz is observed, limited by…
Facial expression recognition (FER) plays a significant role in our daily life. However, annotation ambiguity in the datasets could greatly hinder the performance. In this paper, we address FER task via label distribution learning paradigm,…
A fiber laser is stabilized using a Calcium Fluoride (CaF2) whispering-gallery-mode resonator. It is set up using a semiconductor optical amplifier as a gain medium. The resonator is critically coupled through prisms, and used as a…
A whispering gallery mode resonator with a cavity-made slot filled atomic vapor is demonstrated, in which the chiral symmetry is broken induced by asymmetric backscattering of counter-propagating optical waves in the WGM microcavity. The…
We experimentally and numerically study the potential of photoacoustic-guiding for light focusing through scattering samples via wavefront-shaping and iterative optimization. We experimentally demonstrate that the focusing efficiency on an…
Deep learning has proved successful in many applications but suffers from high computational demands and requires custom accelerators for deployment. Crossbar-based analog in-memory architectures are attractive for acceleration of deep…
Most state-of-the-art self-supervised speaker verification systems rely on a contrastive-based objective function to learn speaker representations from unlabeled speech data. We explore different ways to improve the performance of these…
The evanescent coupling of light between a whispering-gallery-mode bottle microresonator and a sub-wavelength-diameter coupling fiber is actively stabilized by means of a Pound-Drever-Hall technique. We demonstrate the stabilization of a…
Large language models (LLMs) for table-based reasoning often struggle with large tables due to input length limits. We propose ATF (Adaptive Table Filtering Framework), a modular and question-aware filtering pipeline that prunes…
Unsupervised Anomalous Sound Detection (ASD) aims to design a generalizable method that can be used to detect anomalies when only normal sounds are given. In this paper, Anomalous Sound Detection based on Diffusion Models (ASD-Diffusion) is…
Noise-robust automatic speech recognition (ASR) has been commonly addressed by applying speech enhancement (SE) at the waveform level before recognition. However, speech-level enhancement does not always translate into consistent…
Graph anomaly detection (GAD) has garnered increasing attention in recent years, yet remains challenging due to two key factors: (1) label scarcity stemming from the high cost of annotations and (2) homophily disparity at node and class…
Recent advancements in Text-to-Speech (TTS) models, particularly in voice cloning, have intensified the demand for adaptable and efficient deepfake detection methods. As TTS systems continue to evolve, detection models must be able to…
We use the Pound-Drever-Hall (PDH) technique to characterize the frequency stability of a microwave-frequency surface acoustic wave (SAW) resonator-based sensor. The multi-mode acoustic resonator is integrated in a notch geometry with a…
We present experimental results on a Josephson parametric amplifier tailored for readout of ultra-sensitive thermal microwave detectors. In particular, we discuss the impact of fabrication details on the performance. We show that the small…