中文
相关论文

相关论文: SIDAR: Synthetic Image Dataset for Alignment & Res…

200 篇论文

Though many attempts have been made in blind super-resolution to restore low-resolution images with unknown and complex degradations, they are still far from addressing general real-world degraded images. In this work, we extend the…

图像与视频处理 · 电气工程与系统科学 2021-08-18 Xintao Wang , Liangbin Xie , Chao Dong , Ying Shan

Data augmentation is one of the most prevalent tools in deep learning, underpinning many recent advances, including those from classification, generative models, and representation learning. The standard approach to data augmentation…

计算机视觉与模式识别 · 计算机科学 2025-06-12 Brandon Trabucco , Kyle Doherty , Max Gurinas , Ruslan Salakhutdinov

The low dynamic range (LDR) of common cameras fails to capture the rich contrast in natural scenes, resulting in loss of color and details in saturated pixels. Reconstructing the high dynamic range (HDR) of luminance present in the scene…

计算机视觉与模式识别 · 计算机科学 2024-09-24 Sebastian Dille , Chris Careaga , Yağız Aksoy

Due to the difficulty in collecting large-scale and perfectly aligned paired training data for Under-Display Camera (UDC) image restoration, previous methods resort to monitor-based image systems or simulation-based methods, sacrificing the…

计算机视觉与模式识别 · 计算机科学 2023-04-13 Ruicheng Feng , Chongyi Li , Huaijin Chen , Shuai Li , Jinwei Gu , Chen Change Loy

Physical photographs now can be conveniently scanned by smartphones and stored forever as a digital version, yet the scanned photos are not restored well. One solution is to train a supervised deep neural network on many digital photos and…

计算机视觉与模式识别 · 计算机科学 2021-08-19 Man M. Ho , Jinjia Zhou

With the increase of computing power, machine learning models in medical imaging have been introduced to help in rending medical diagnosis and inspection, like hemophilia, a rare disorder in which blood cannot clot normally. Often, one of…

计算机视觉与模式识别 · 计算机科学 2024-09-19 Qianyu Fan

Deep learning-based image retrieval has been emphasized in computer vision. Representation embedding extracted by deep neural networks (DNNs) not only aims at containing semantic information of the image, but also can manage large-scale…

计算机视觉与模式识别 · 计算机科学 2021-09-22 Seonho Park , Maciej Rysz , Kathleen M. Dipple , Panos M. Pardalos

Synthetic Aperture Radar (SAR) and optical image registration is essential for remote sensing data fusion, with applications in military reconnaissance, environmental monitoring, and disaster management. However, challenges arise from…

计算机视觉与模式识别 · 计算机科学 2025-02-04 Wenfei Zhang , Ruipeng Zhao , Yongxiang Yao , Yi Wan , Peihao Wu , Jiayuan Li , Yansheng Li , Yongjun Zhang

In recent years, computer vision has transformed fields such as medical imaging, object recognition, and geospatial analytics. One of the fundamental tasks in computer vision is semantic image segmentation, which is vital for precise object…

计算机视觉与模式识别 · 计算机科学 2023-11-09 Dinar Sharafutdinov , Stanislav Kuskov , Saian Protasov , Alexey Voropaev

Fusion-based hyperspectral image (HSI) super-resolution has become increasingly prevalent for its capability to integrate high-frequency spatial information from the paired high-resolution (HR) RGB reference image. However, most of the…

计算机视觉与模式识别 · 计算机科学 2023-02-14 Zeqiang Lai , Ying Fu , Jun Zhang

Real-world point cloud datasets have made significant contributions to the development of LiDAR-based perception technologies, such as object segmentation for autonomous driving. However, due to the limited number of instances in some rare…

计算机视觉与模式识别 · 计算机科学 2025-10-14 Shutong Lin , Zhengkang Xiang , Jianzhong Qi , Kourosh Khoshelham

Class imbalance is a persistent challenge in visual recognition, particularly in safety-critical domains where collecting positive examples is expensive and rare events are inherently underrepresented. We propose a lightweight synthetic…

计算机视觉与模式识别 · 计算机科学 2026-05-13 Daniil Dushenev , Nazariy Karpov , Daniil Zinovjev , Alexander Gorin , Konstantin Kulikov

High Dynamic Range (HDR) imaging aims to reproduce the wide range of brightness levels present in natural scenes, which the human visual system can perceive but conventional digital cameras often fail to capture due to their limited dynamic…

图像与视频处理 · 电气工程与系统科学 2025-10-28 Kumbha Nagaswetha

Recent work has shown that data augmentation has the potential to significantly improve the generalization of deep learning models. Recently, automated augmentation strategies have led to state-of-the-art results in image classification and…

计算机视觉与模式识别 · 计算机科学 2019-11-15 Ekin D. Cubuk , Barret Zoph , Jonathon Shlens , Quoc V. Le

We introduce a multi-scale Image Super Resolution (ISR) method building on recent advances in Visual Auto-Regressive (VAR) modeling. VAR models break image tokenization into additive, gradually increasing scales, using Residual Quantization…

计算机视觉与模式识别 · 计算机科学 2026-05-15 Isma Hadji , Enrique Sanchez , Adrian Bulat , Brais Martinez , Georgios Tzimiropoulos

We present a model that generates natural language descriptions of images and their regions. Our approach leverages datasets of images and their sentence descriptions to learn about the inter-modal correspondences between language and…

计算机视觉与模式识别 · 计算机科学 2015-04-15 Andrej Karpathy , Li Fei-Fei

Hyperspectral image (HSI) restoration aims at recovering clean images from degraded observations and plays a vital role in downstream tasks. Existing model-based methods have limitations in accurately modeling the complex image…

计算机视觉与模式识别 · 计算机科学 2024-02-27 Li Pang , Xiangyu Rui , Long Cui , Hongzhong Wang , Deyu Meng , Xiangyong Cao

Tomography is an imaging technique that works by reconstructing a scene from acquired data in the form of line integrals of the imaging domain. A fundamental underlying assumption in the reconstruction procedure is the precise alignment of…

数值分析 · 数学 2018-07-13 Toby Sanders

Depth estimation is an essential component in understanding the 3D geometry of a scene, with numerous applications in urban and indoor settings. These scenes are characterized by a prevalence of human made structures, which in most of the…

计算机视觉与模式识别 · 计算机科学 2020-09-03 Mattia Rossi , Mireille El Gheche , Andreas Kuhn , Pascal Frossard

3D occupancy prediction aims to infer dense, voxel-wise scene semantics from sensor observations, where the 2D-to-3D view transformation serves as a crucial step in bridging image features and volumetric representations. Most previous…

计算机视觉与模式识别 · 计算机科学 2026-05-12 Yuan Wu , Zhiqiang Yan , Jiawei Lian , Zhengxue Wang , Jian Yang