中文
相关论文

相关论文: Amanous: Distribution-Switching for Superhuman Pia…

200 篇论文

In recent years, advancements in neural network designs and the availability of large-scale labeled datasets have led to significant improvements in the accuracy of piano transcription models. However, most previous work focused on…

音频与语音处理 · 电气工程与系统科学 2024-04-11 Taegyun Kwon , Dasaem Jeong , Juhan Nam

This paper investigates score-based diffusion models when the underlying target distribution is concentrated on or near low-dimensional manifolds within the higher-dimensional space in which they formally reside, a common characteristic of…

机器学习 · 计算机科学 2025-01-03 Gen Li , Yuling Yan

Anomaly detection in medical imaging plays a crucial role in identifying pathological regions across various imaging modalities, such as brain MRI, liver CT, and carotid ultrasound (US). However, training fully supervised segmentation…

图像与视频处理 · 电气工程与系统科学 2025-07-29 Yuan Bi , Lucie Huang , Ricarda Clarenbach , Reza Ghotbi , Angelos Karlas , Nassir Navab , Zhongliang Jiang

Algorithms increasingly operate within complex physical, social, and engineering systems where they are exposed to disturbances, noise, and interconnections with other dynamical systems. This article extends known convergence guarantees of…

机器学习 · 计算机科学 2025-12-22 Guner Dilsad Er , Sebastian Trimpe , Michael Muehlebach

Large high-dimensional datasets are becoming more and more popular in an increasing number of research areas. Processing the high dimensional data incurs a high computational cost and is inherently inefficient since many of the values that…

计算机视觉与模式识别 · 计算机科学 2013-05-01 Alon Schclar

Diffusion models (DMs) have established themselves as the state-of-the-art generative modeling approach in the visual domain and beyond. A crucial drawback of DMs is their slow sampling speed, relying on many sequential function evaluations…

计算机视觉与模式识别 · 计算机科学 2024-04-24 Amirmojtaba Sabour , Sanja Fidler , Karsten Kreis

The remarkable capabilities of pretrained image diffusion models have been utilized not only for generating fixed-size images but also for creating panoramas. However, naive stitching of multiple images often results in visible seams.…

计算机视觉与模式识别 · 计算机科学 2023-10-31 Yuseung Lee , Kunho Kim , Hyunjin Kim , Minhyuk Sung

Acoustic transparency is the capability of a medium to transmit mechanical waves to adjacent media, without scattering. This characteristic can be achieved by carefully engineering the acoustic impedance of the medium -- a combination of…

应用物理 · 物理学 2021-06-22 Sai Sharan Injeti , Paolo Celli , Kaushik Bhattacharya , Chiara Daraio

Change detection involves segmenting sequential data such that observations in the same segment share some desired properties. Multivariate change detection continues to be a challenging problem due to the variety of ways change points can…

统计方法学 · 统计学 2018-10-16 Wenyu Zhang , Daniel Gilbert , David Matteson

This paper presents a sensory fusion neuromorphic dataset collected with precise temporal synchronization using a set of Address-Event-Representation sensors and tools. The target application is the lip reading of several keywords for…

As diffusion-based deep generative models gain prevalence, researchers are actively investigating their potential applications across various domains, including music synthesis and style alteration. Within this work, we are interested in…

音频与语音处理 · 电气工程与系统科学 2024-09-25 Teysir Baoueb , Xiaoyu Bie , Hicham Janati , Gael Richard

Gradient-type distributed optimization methods have blossomed into one of the most important tools for solving a minimization learning task over a networked agent system. However, only one gradient update per iteration is difficult to…

最优化与控制 · 数学 2024-03-06 Mou Wu , Haibin Liao , Zhengtao Ding , Yonggang Xiao

By decomposing the image formation process into a sequential application of denoising autoencoders, diffusion models (DMs) achieve state-of-the-art synthesis results on image data and beyond. Additionally, their formulation allows for a…

计算机视觉与模式识别 · 计算机科学 2022-04-14 Robin Rombach , Andreas Blattmann , Dominik Lorenz , Patrick Esser , Björn Ommer

In this paper we study a special class of systems: time-invariant control systems that satisfy the matching condition for which no bounds for the disturbance and the unknown parameters are known. For this class of systems, we provide a…

最优化与控制 · 数学 2023-11-15 Iasson Karafyllis , Miroslav Krstic

Diffusion and flow-matching models have revolutionized automatic text-to-audio generation in recent times. These models are increasingly capable of generating high quality and faithful audio outputs capturing to speech and acoustic events.…

Learned models and policies can generalize effectively when evaluated within the distribution of the training data, but can produce unpredictable and erroneous outputs on out-of-distribution inputs. In order to avoid distribution shift when…

机器学习 · 计算机科学 2022-06-22 Katie Kang , Paula Gradu , Jason Choi , Michael Janner , Claire Tomlin , Sergey Levine

In the era of large-scale pre-trained models, effectively adapting general knowledge to specific affective computing tasks remains a challenge, particularly regarding computational efficiency and multimodal heterogeneity. While…

人工智能 · 计算机科学 2026-03-20 Yan Li , Yifei Xing , Xiangyuan Lan , Xin Li , Haifeng Chen , Dongmei Jiang

Computational phantoms are widely used in medical imaging research, yet current systems to generate controlled, clinically meaningful anatomical variations remain limited. We present AbdomenGen, a sequential volume-conditioned diffusion…

计算机视觉与模式识别 · 计算机科学 2026-04-15 Yubraj Bhandari , Lavsen Dahal , Paul Segars , Joseph Y. Lo

While deep generative models have become the leading methods for algorithmic composition, it remains a challenging problem to control the generation process because the latent variables of most deep-learning models lack good…

声音 · 计算机科学 2020-08-18 Ziyu Wang , Dingsu Wang , Yixiao Zhang , Gus Xia

To develop a sound-monitoring system for machines, a method for detecting anomalous sound under domain shifts is proposed. A domain shift occurs when a machine's physical parameters change. Because a domain shift changes the distribution of…

音频与语音处理 · 电气工程与系统科学 2021-11-15 Kota Dohi , Takashi Endo , Yohei Kawaguchi