中文
相关论文

相关论文: RAW-Adapter: Adapting Pre-trained Visual Model to …

200 篇论文

White balance (WB) is a key step in the image signal processor (ISP) pipeline that mitigates color casts caused by varying illumination and restores the scene's true colors. Currently, sRGB-based WB editing for post-ISP WB correction is…

计算机视觉与模式识别 · 计算机科学 2025-12-12 Yang Cheng , Ziteng Cui , Shenghan Su , Lin Gu , Zenghui Zhang

Robotic Perception in diverse domains such as low-light scenarios, where new modalities like thermal imaging and specialized night-vision sensors are increasingly employed, remains a challenge. Largely, this is due to the limited…

计算机视觉与模式识别 · 计算机科学 2023-06-21 Anirudha Ramesh , Anurag Ghosh , Christoph Mertz , Jeff Schneider

Image reconstruction from corrupted images is crucial across many domains. Most reconstruction networks are trained on post-ISP sRGB images, even though the image-signal-processing pipeline irreversibly mixes colors, clips dynamic range,…

计算机视觉与模式识别 · 计算机科学 2025-12-30 Nate Rothschild , Moshe Kimhi , Avi Mendelson , Chaim Baskin

RAW files are the initial measurement of scene radiance widely used in most cameras, and the ubiquitously-used RGB images are converted from RAW data through Image Signal Processing (ISP) pipelines. Nowadays, digital images are risky of…

计算机视觉与模式识别 · 计算机科学 2023-08-01 Xiaoxiao Hu , Qichao Ying , Zhenxing Qian , Sheng Li , Xinpeng Zhang

The success of deep denoisers on real-world color photographs usually relies on the modeling of sensor noise and in-camera signal processing (ISP) pipeline. Performance drop will inevitably happen when the sensor and ISP pipeline of test…

计算机视觉与模式识别 · 计算机科学 2021-03-19 Yue Cao , Xiaohe Wu , Shuran Qi , Xiao Liu , Zhongqin Wu , Wangmeng Zuo

Transformer-based models have transformed the landscape of natural language processing (NLP) and are increasingly applied to computer vision tasks with remarkable success. These models, renowned for their ability to capture long-range…

计算机视觉与模式识别 · 计算机科学 2024-08-28 Gracile Astlin Pereira , Muhammad Hussain

Most visual models are designed for sRGB images, yet RAW data offers significant advantages for object detection by preserving sensor information before ISP processing. This enables improved detection accuracy and more efficient hardware…

计算机视觉与模式识别 · 计算机科学 2025-11-18 Haiyang Xie , Xi Shen , Shihua Huang , Qirui Wang , Zheng Wang

We introduce the BIR-Adapter, a parameter-efficient diffusion adapter for blind image restoration. Diffusion-based restoration methods have demonstrated promising performance in addressing this fundamental problem in computer vision,…

计算机视觉与模式识别 · 计算机科学 2026-04-28 Cem Eteke , Alexander Griessel , Wolfgang Kellerer , Eckehard Steinbach

Modern cameras typically offer two types of image states: a minimally processed linear raw RGB image representing the raw sensor data, and a highly-processed non-linear image state, such as the sRGB state. The CIE-XYZ color space is a…

图像与视频处理 · 电气工程与系统科学 2024-05-22 Shir Barzel , Moshe Salhov , Ofir Lindenbaum , Amir Averbuch

In order to deploy current computer vision (CV) models on resource-constrained low-power devices, recent works have proposed in-sensor and in-pixel computing approaches that try to partly/fully bypass the image signal processor (ISP) and…

计算机视觉与模式识别 · 计算机科学 2022-10-12 Gourav Datta , Zeyu Liu , Zihan Yin , Linyu Sun , Akhilesh R. Jaiswal , Peter A. Beerel

Raw images preserve linear sensor measurements and high bit-depth information crucial for advanced vision tasks and photography applications, yet their storage remains challenging due to large file sizes, varying bit depths, and…

计算机视觉与模式识别 · 计算机科学 2026-03-31 Chunhang Zheng , Tongda Xu , Mingli Xie , Yan Wang , Dou Li

Advancements in deep learning have ignited an explosion of research on efficient hardware for embedded computer vision. Hardware vision acceleration, however, does not address the cost of capturing and processing the image data that feeds…

计算机视觉与模式识别 · 计算机科学 2017-08-03 Mark Buckler , Suren Jayasuriya , Adrian Sampson

With the advent of large-scale pre-trained models, interest in adapting and exploiting them for continual learning scenarios has grown. In this paper, we propose an approach to exploiting pre-trained vision-language models (e.g. CLIP) that…

计算机视觉与模式识别 · 计算机科学 2023-11-01 Xialei Liu , Xusheng Cao , Haori Lu , Jia-wen Xiao , Andrew D. Bagdanov , Ming-Ming Cheng

Deep artificial neural networks, trained with labeled data sets are widely used in numerous vision and robotics applications today. In terms of AI, these are called reflex models, referring to the fact that they do not self-evolve or…

计算机视觉与模式识别 · 计算机科学 2020-02-20 Hai Xiao , Jin Shang , Mengyuan Huang

Downsampling is widely adopted to achieve a good trade-off between accuracy and latency for visual recognition. Unfortunately, the commonly used pooling layers are not learned, and thus cannot preserve important information. As another…

计算机视觉与模式识别 · 计算机科学 2022-07-26 Ho Man Kwan , Shenghui Song

Advances in high dynamic range (HDR) lighting estimation from a single image have opened new possibilities for augmented reality (AR) applications. Predicting complex lighting environments from a single input image allows for the realistic…

计算机视觉与模式识别 · 计算机科学 2025-09-24 Zitian Zhang , Joshua Urban Davis , Jeanne Phuong Anh Vu , Jiangtao Kuang , Jean-François Lalonde

Existing reference (RF)-based super-resolution (SR) models try to improve perceptual quality in SR under the assumption of the availability of high-resolution RF images paired with low-resolution (LR) inputs at testing. As the RF images…

计算机视觉与模式识别 · 计算机科学 2021-04-07 Mohammad Saeed Rad , Thomas Yu , Behzad Bozorgtabar , Jean-Philippe Thiran

In this paper, we delve into the concept of interpretable image enhancement, a technique that enhances image quality by adjusting filter parameters with easily understandable names such as "Exposure" and "Contrast". Unlike using predefined…

计算机视觉与模式识别 · 计算机科学 2024-08-21 Satoshi Kosugi

Conventional image signal processing (ISP) frameworks are designed to reconstruct an RGB image from a single raw measurement. As multi-camera systems become increasingly popular these days, it is worth exploring improvements in ISP…

图像与视频处理 · 电气工程与系统科学 2022-11-16 Ahmad Bin Rabiah , Qi Guo

Large-scale contrastive vision-language pre-training has shown significant progress in visual representation learning. Unlike traditional visual systems trained by a fixed set of discrete labels, a new paradigm was introduced in…

计算机视觉与模式识别 · 计算机科学 2025-03-26 Peng Gao , Shijie Geng , Renrui Zhang , Teli Ma , Rongyao Fang , Yongfeng Zhang , Hongsheng Li , Yu Qiao