中文
相关论文

相关论文: LaRE$^2$: Latent Reconstruction Error Based Method…

200 篇论文

Current deep learning approaches in computer vision primarily focus on RGB data sacrificing information. In contrast, RAW images offer richer representation, which is crucial for precise recognition, particularly in challenging conditions…

计算机视觉与模式识别 · 计算机科学 2024-11-21 Christoph Reinders , Radu Berdan , Beril Besbinar , Junji Otsuka , Daisuke Iso

While learned image compression (LIC) focuses on efficient data transmission, generative image compression (GIC) extends this framework by integrating generative modeling to produce photo-realistic reconstructed images. In this paper, we…

图像与视频处理 · 电气工程与系统科学 2025-05-28 Minghao Han , Weiyi You , Jinhua Zhang , Leheng Zhang , Ce Zhu , Shuhang Gu

Diffusion models have significantly advanced generative AI, but they encounter difficulties when generating complex combinations of multiple objects. As the final result heavily depends on the initial seed, accurately ensuring the desired…

计算机视觉与模式识别 · 计算机科学 2024-09-18 Federico Betti , Lorenzo Baraldi , Lorenzo Baraldi , Rita Cucchiara , Nicu Sebe

Existing state-of-the-art AI-Generated image detection methods mostly consider extracting low-level information from RGB images to help improve the generalization of AI-Generated image detection, such as noise patterns. However, these…

计算机视觉与模式识别 · 计算机科学 2025-07-18 Ziyin Zhou , Ke Sun , Zhongxi Chen , Xianming Lin , Yunpeng Luo , Ke Yan , Shouhong Ding , Xiaoshuai Sun

Latent diffusion models with Transformer architectures excel at generating high-fidelity images. However, recent studies reveal an optimization dilemma in this two-stage design: while increasing the per-token feature dimension in visual…

计算机视觉与模式识别 · 计算机科学 2025-03-11 Jingfeng Yao , Bin Yang , Xinggang Wang

Late gadolinium enhancement (LGE) imaging is the clinical standard for myocardial scar assessment, but limited annotated datasets hinder the development of automated segmentation methods. We propose a novel framework that synthesises both…

计算机视觉与模式识别 · 计算机科学 2026-02-13 Soufiane Ben Haddou , Laura Alvarez-Florez , Erik J. Bekkers , Fleur V. Y. Tjong , Ahmad S. Amin , Connie R. Bezzina , Ivana Išgum

Absolute camera pose estimation is usually addressed by sequentially solving two distinct subproblems: First a feature matching problem that seeks to establish putative 2D-3D correspondences, and then a Perspective-n-Point problem that…

计算机视觉与模式识别 · 计算机科学 2021-03-15 Hugo Germain , Vincent Lepetit , Guillaume Bourmaud

The rapid rise of generative models has yielded synthetic images of striking realism, blurring the line between real and fake content. As novel models proliferate, detectors must go beyond mere fake identification to robustly generalise…

计算机视觉与模式识别 · 计算机科学 2026-03-11 Simone Bonechi , Paolo Andreini , Barbara Toniella Corradini

Artificial Intelligence-Generated Content (AIGC) has made significant strides, with high-resolution text-to-image (T2I) generation becoming increasingly critical for improving users' Quality of Experience (QoE). Although…

计算机视觉与模式识别 · 计算机科学 2026-01-22 Chongbin Yi , Yuxin Liang , Ziqi Zhou , Peng Yang

Low-light image enhancement (LLIE) has achieved promising performance by employing conditional diffusion models. Despite the success of some conditional methods, previous methods may neglect the importance of a sufficient formulation of…

计算机视觉与模式识别 · 计算机科学 2024-07-30 Yuhui Wu , Guoqing Wang , Zhiwen Wang , Yang Yang , Tianyu Li , Malu Zhang , Chongyi Li , Heng Tao Shen

Capturing High Dynamic Range (HDR) scenery using 8-bit cameras often suffers from over-/underexposure, loss of fine details due to low bit-depth compression, skewed color distributions, and strong noise in dark areas. Traditional LDR image…

图像与视频处理 · 电气工程与系统科学 2024-06-14 Baiang Li , Sizhuo Ma , Yanhong Zeng , Xiaogang Xu , Youqing Fang , Zhao Zhang , Jian Wang , Kai Chen

Dramatic advances in the quality of the latent diffusion models (LDMs) also led to the malicious use of AI-generated images. While current AI-generated image detection methods assume the availability of real/AI-generated images for…

计算机视觉与模式识别 · 计算机科学 2026-04-14 Sungik Choi , Hankook Lee , Jaehoon Lee , Seunghyun Kim , Stanley Jungkyu Choi , Moontae Lee

In neural decoding research, one of the most intriguing topics is the reconstruction of perceived natural images based on fMRI signals. Previous studies have succeeded in re-creating different aspects of the visuals, such as low-level…

计算机视觉与模式识别 · 计算机科学 2023-06-22 Furkan Ozcelik , Rufin VanRullen

Autoregressive (AR) image generators offer a language-model-friendly approach to image generation by predicting discrete image tokens in a causal sequence. However, unlike diffusion models, AR models lack a mechanism to refine previous…

计算机视觉与模式识别 · 计算机科学 2026-01-29 Cheng Cheng , Lin Song , Di An , Yicheng Xiao , Xuchong Zhang , Hongbin Sun , Ying Shan

Latent diffusion models have established a new state-of-the-art in high-resolution visual generation. Integrating Vision Foundation Model priors improves generative efficiency, yet existing latent designs remain largely heuristic. These…

计算机视觉与模式识别 · 计算机科学 2026-03-13 Hangyu Liu , Jianyong Wang , Yutao Sun

Self-supervised representation learning has gained increasing attention for strong generalization ability without relying on paired datasets. However, it has not been explored sufficiently for facial representation. Self-supervised facial…

计算机视觉与模式识别 · 计算机科学 2024-05-24 Ruian He , Zhen Xing , Weimin Tan , Bo Yan

Concept erasure aims to suppress sensitive content in diffusion models, but recent studies show that erased concepts can still be reawakened, revealing vulnerabilities in erasure methods. Existing reawakening methods mainly rely on…

计算机视觉与模式识别 · 计算机科学 2026-05-19 Mengyu Sun , Ziyuan Yang , Andrew Beng Jin Teoh , Junxu Liu , Haibo Hu , Yi Zhang

Text-based image editing, powered by generative diffusion models, lets users modify images through natural-language prompts and has dramatically simplified traditional workflows. Despite these advances, current methods still suffer from a…

计算机视觉与模式识别 · 计算机科学 2025-08-26 Sunung Mun , Jinhwan Nam , Sunghyun Cho , Jungseul Ok

In this work, we consider the image super-resolution (SR) problem. The main challenge of image SR is to recover high-frequency details of a low-resolution (LR) image that are important for human perception. To address this essentially…

计算机视觉与模式识别 · 计算机科学 2017-11-22 Wenhan Yang , Jiashi Feng , Jianchao Yang , Fang Zhao , Jiaying Liu , Zongming Guo , Shuicheng Yan

Single-image reflection removal is a highly ill-posed problem, where existing methods struggle to reason about the composition of corrupted regions, causing them to fail at recovery and generalization in the wild. This work reframes an…

计算机视觉与模式识别 · 计算机科学 2025-12-09 Mingjia Li , Jin Hu , Hainuo Wang , Qiming Hu , Jiarui Wang , Xiaojie Guo