中文
相关论文

相关论文: MetaISP -- Exploiting Global Scene Structure for A…

200 篇论文

Night Photography Rendering (NPR) poses a significant challenge due to the extreme contrast between dark and illuminated areas in scenes, stemming from concurrent capture of severely dark regions alongside intense point light sources.…

计算机视觉与模式识别 · 计算机科学 2026-05-01 Furkan Kınlı

Super-resolution is a fundamental problem in computer vision which aims to overcome the spatial limitation of camera sensors. While significant progress has been made in single image super-resolution, most algorithms only perform well on…

计算机视觉与模式识别 · 计算机科学 2021-02-03 Xiangyu Xu , Yongrui Ma , Wenxiu Sun , Ming-Hsuan Yang

Scene Text Recognition requires modeling visual structures that evolve from coarse layouts to fine-grained character strokes. Training such models relies on large amounts of annotated data. Recent self-supervised approaches, such as Masked…

计算机视觉与模式识别 · 计算机科学 2026-05-15 Zhuohao Chen , Zeng Li , Yifei Zhang , Chang Liu , Yu Zhou

Compared to DSLR cameras, smartphone cameras have smaller sensors, which limits their spatial resolution; smaller apertures, which limits their light gathering ability; and smaller pixels, which reduces their signal-to noise ratio. The use…

计算机视觉与模式识别 · 计算机科学 2021-02-18 Bartlomiej Wronski , Ignacio Garcia-Dorado , Manfred Ernst , Damien Kelly , Michael Krainin , Chia-Kai Liang , Marc Levoy , Peyman Milanfar

We exploit human color metamers to send light-modulated messages less visible to the human eye, but recoverable by cameras. These messages are a key component to camera-display messaging, such as handheld smartphones capturing information…

计算机视觉与模式识别 · 计算机科学 2016-04-07 Eric Wengrowski , Kristin Dana , Marco Gruteser , Narayan Mandayam

Multi-view 3D reconstruction methods remain highly sensitive to photometric inconsistencies arising from camera optical characteristics and variations in image signal processing (ISP). Existing mitigation strategies such as per-frame latent…

计算机视觉与模式识别 · 计算机科学 2026-04-08 Isaac Deutsch , Nicolas Moënne-Loccoz , Gavriel State , Zan Gojcic

Image reconstruction from corrupted images is crucial across many domains. Most reconstruction networks are trained on post-ISP sRGB images, even though the image-signal-processing pipeline irreversibly mixes colors, clips dynamic range,…

计算机视觉与模式识别 · 计算机科学 2025-12-30 Nate Rothschild , Moshe Kimhi , Avi Mendelson , Chaim Baskin

Digital zoom on smartphones relies on learning-based super-resolution (SR) models that operate on RAW sensor images, but obtaining sensor-specific training data is challenging due to the lack of ground-truth images. Synthetic data…

计算机视觉与模式识别 · 计算机科学 2026-03-16 Ali Mosleh , Faraz Ali , Fengjia Zhang , Stavros Tsogkas , Junyong Lee , Alex Levinshtein , Michael S. Brown

Infrared image super-resolution (IISR) under real-world conditions is a practically significant yet rarely addressed task. Pioneering works are often trained and evaluated on simulated datasets or neglect the intrinsic differences between…

计算机视觉与模式识别 · 计算机科学 2026-03-06 Yang Zou , Jun Ma , Zhidong Jiao , Xingyuan Li , Zhiying Jiang , Jinyuan Liu

Digital Signal Processing (DSP) and Digital Image Processing (DIP) with Machine Learning (ML) and Deep Learning (DL) are popular research areas in Computer Vision and related fields. We highlight transformative applications in image…

Existing remote sensing change detection methods are heavily affected by seasonal variation. Since vegetation colors are different between winter and summer, such variations are inclined to be falsely detected as changes. In this letter, we…

计算机视觉与模式识别 · 计算机科学 2022-02-16 Tiange Zhang , Feng Gao , Junyu Dong , Qian Du

Semantic segmentation is a process of partitioning an image into multiple segments for recognizing humans and objects, which can be widely applied in scenarios such as healthcare and safety monitoring. To avoid privacy violation, using RF…

信号处理 · 电气工程与系统科学 2021-10-07 Jingzhi Hu , Hongliang Zhang , Kaigui Bian , Zhu Han , H. Vincent Poor , Lingyang Song

This work tackles the challenging task of achieving real-time novel view synthesis for reflective surfaces across various scenes. Existing real-time rendering methods, especially those based on meshes, often have subpar performance in…

计算机视觉与模式识别 · 计算机科学 2024-08-16 Chaojie Ji , Yufeng Li , Yiyi Liao

We introduce Synscapes -- a synthetic dataset for street scene parsing created using photorealistic rendering techniques, and show state-of-the-art results for training and validation as well as new types of analysis. We study the behavior…

计算机视觉与模式识别 · 计算机科学 2018-10-23 Magnus Wrenninge , Jonas Unger

Self-supervised learning (SSL) methods targeting scene images have seen a rapid growth recently, and they mostly rely on either a dedicated dense matching mechanism or a costly unsupervised object discovery module. This paper shows that…

计算机视觉与模式识别 · 计算机科学 2023-10-02 Ke Zhu , Minghao Fu , Jianxin Wu

Scene classification is a fundamental perception task for environmental understanding in today's robotics. In this paper, we have attempted to exploit the use of popular machine learning technique of deep learning to enhance scene…

计算机视觉与模式识别 · 计算机科学 2015-09-23 Yiyi Liao , Sarath Kodagoda , Yue Wang , Lei Shi , Yong Liu

Advances in 3D reconstruction using neural rendering have enabled high-quality 3D capture. However, they often fail when the input imagery is corrupted by motion blur, due to fast motion of the camera or the objects in the scene. This work…

图像与视频处理 · 电气工程与系统科学 2025-06-30 Sai Sri Teja , Sreevidya Chintalapati , Vinayak Gupta , Mukund Varma T , Haejoon Lee , Aswin Sankaranarayanan , Kaushik Mitra

Physical photographs now can be conveniently scanned by smartphones and stored forever as a digital version, yet the scanned photos are not restored well. One solution is to train a supervised deep neural network on many digital photos and…

计算机视觉与模式识别 · 计算机科学 2021-08-19 Man M. Ho , Jinjia Zhou

Text-to-image (T2I) models have ushered in a new era of real-world image super-resolution (Real-ISR) due to their rich internal implicit knowledge for multimodal learning. Although bringing high-level semantic priors and dense pixel…

计算机视觉与模式识别 · 计算机科学 2025-12-03 Xinrui Li , Jinrong Zhang , Jianlong Wu , Chong Chen , Liqiang Nie , Zhouchen Lin

The spectral response of a digital camera defines the mapping between scene radiance and pixel intensity. Despite its critical importance, there is currently no comprehensive model that considers the end-to-end interaction between light…

图像与视频处理 · 电气工程与系统科学 2025-07-08 Sanush K Abeysekera , Ye Chow Kuang , Melanie Po-Leen Ooi