中文
相关论文

相关论文: Omegance: A Single Parameter for Various Granulari…

200 篇论文

Diffusion models represent a class of generative models that produce data by denoising a sample corrupted by white noise. Despite the success of diffusion models in computer vision, audio synthesis, and point cloud generation, so far they…

统计力学 · 物理学 2025-01-17 Kanta Masuki , Yuto Ashida

Despite significant progress in text-to-image diffusion models, achieving precise spatial control over generated outputs remains challenging. ControlNet addresses this by introducing an auxiliary conditioning module, while ControlNet++…

计算机视觉与模式识别 · 计算机科学 2025-07-04 Nina Konovalova , Maxim Nikolaev , Andrey Kuznetsov , Aibek Alanov

While diffusion models achieve state-of-the-art generation quality, they still suffer from computationally expensive sampling. Recent works address this issue with gradient-based optimization methods that distill a few-step ODE diffusion…

The biomedical imaging world is notorious for working with small amounts of data, frustrating state-of-the-art efforts in the computer vision and deep learning worlds. With large datasets, it is easier to make progress we have seen from the…

计算机视觉与模式识别 · 计算机科学 2022-12-08 Manuel Serna-Aguilera , Khoa Luu , Nathaniel Harris , Min Zou

Spectral synthesis is a powerful tool with which to find the fundamental parameters of stars. Models are usually restricted to single values of temperature and gravity, and assume spherical symmetry. This approximation breaks down for…

太阳与恒星天体物理 · 物理学 2024-06-27 Benjamin Montesinos

In this paper, we present the Directly Denoising Diffusion Model (DDDM): a simple and generic approach for generating realistic images with few-step sampling, while multistep sampling is still preserved for better performance. DDDMs require…

计算机视觉与模式识别 · 计算机科学 2024-06-03 Dan Zhang , Jingjing Wang , Feng Luo

We introduce a novel state-space architecture for diffusion models, effectively harnessing spatial and frequency information to enhance the inductive bias towards local features in input images for image generation tasks. While state-space…

计算机视觉与模式识别 · 计算机科学 2025-04-14 Hao Phung , Quan Dao , Trung Dao , Hoang Phan , Dimitris Metaxas , Anh Tran

Decomposing geometry, materials and lighting from a set of images, namely inverse rendering, has been a long-standing problem in computer vision and graphics. Recent advances in neural rendering enable photo-realistic and plausible inverse…

计算机视觉与模式识别 · 计算机科学 2025-03-04 Silong Yong , Venkata Nagarjun Pudureddiyur Manivannan , Bernhard Kerbl , Zifu Wan , Simon Stepputtis , Katia Sycara , Yaqi Xie

This paper studies a diffusion-based framework to address the low-light image enhancement problem. To harness the capabilities of diffusion models, we delve into this intricate process and advocate for the regularization of its inherent…

计算机视觉与模式识别 · 计算机科学 2023-10-30 Jinhui Hou , Zhiyu Zhu , Junhui Hou , Hui Liu , Huanqiang Zeng , Hui Yuan

Diffusion models have been verified to be effective in generating complex distributions from natural images to motion trajectories. Recent diffusion-based methods show impressive performance in 3D robotic manipulation tasks, whereas they…

机器人学 · 计算机科学 2025-09-09 Guanxing Lu , Zifeng Gao , Tianxing Chen , Wenxun Dai , Ziwei Wang , Wenbo Ding , Yansong Tang

Efficiently compiling quantum operations remains a major bottleneck in scaling quantum computing. Today's state-of-the-art methods achieve low compilation error by combining search algorithms with gradient-based parameter optimization, but…

量子物理 · 物理学 2026-05-12 Florian Fürrutter , Zohim Chandani , Ikko Hamamura , Hans J. Briegel , Gorka Muñoz-Gil

Diffusion models, emerging as powerful deep generative tools, excel in various applications. They operate through a two-steps process: introducing noise into training samples and then employing a model to convert random noise into new…

计算机视觉与模式识别 · 计算机科学 2026-02-13 Huijie Zhang , Yifu Lu , Ismail Alkhouri , Saiprasad Ravishankar , Dogyoon Song , Qing Qu

Video frame prediction extrapolates future frames from previous frames, but suffers from prediction errors in dynamic scenes due to the lack of information about the next frame. Event cameras address this limitation by capturing per-pixel…

计算机视觉与模式识别 · 计算机科学 2025-12-22 Jiyun Kong , Jun-Hyuk Kim , Jong-Seok Lee

Adapting large-scale pre-trained generative models in a parameter-efficient manner is gaining traction. Traditional methods like low rank adaptation achieve parameter efficiency by imposing constraints but may not be optimal for tasks…

计算机视觉与模式识别 · 计算机科学 2024-06-03 Xinxi Zhang , Song Wen , Ligong Han , Felix Juefei-Xu , Akash Srivastava , Junzhou Huang , Hao Wang , Molei Tao , Dimitris N. Metaxas

Text-to-image diffusion generative models can generate high quality images at the cost of tedious prompt engineering. Controllability can be improved by introducing layout conditioning, however existing methods lack layout editing ability…

计算机视觉与模式识别 · 计算机科学 2024-11-19 Alessandro Fontanella , Petru-Daniel Tudosiu , Yongxin Yang , Shifeng Zhang , Sarah Parisot

We present a novel particle management method using the Characteristic Mapping framework. In the context of explicit evolution of parametrized curves and surfaces, the surface distribution of marker points created from sampling the…

数值分析 · 数学 2023-02-21 Xi-Yuan Yin , Linan Chen , Jean-Christophe Nave

Algorithms for generating random numbers that follow a gamma distribution with shape parameter less than unity are proposed. Acceptance-rejection algorithms are developed, based on the generalized exponential distribution. The squeeze…

统计计算 · 统计学 2024-11-18 Seiji Zenitani

We study the single-species diffusion-annihilation process with a time-dependent reaction rate, lambda(t)=lambda_0 t^-omega. Scaling arguments show that there is a critical value of the decay exponent omega_c(d) separating a…

统计力学 · 物理学 2007-05-23 L. Turban

Diffusion models have recently achieved great success in the synthesis of high-quality images and videos. However, the existing denoising techniques in diffusion models are commonly based on step-by-step noise predictions, which suffers…

计算机视觉与模式识别 · 计算机科学 2024-10-15 Hancheng Ye , Jiakang Yuan , Renqiu Xia , Xiangchao Yan , Tao Chen , Junchi Yan , Botian Shi , Bo Zhang

Recent unified models such as Bagel demonstrate that paired image-edit data can effectively align multiple visual tasks within a single diffusion transformer. However, these models remain limited to single-condition inputs and lack the…

计算机视觉与模式识别 · 计算机科学 2026-02-10 Xiaoyan Zhang , Zechen Bai , Haofan Wang , Yiren Song