English
Related papers

Related papers: MERIT: Multi-domain Efficient RAW Image Translatio…

200 papers

Current deep learning approaches in computer vision primarily focus on RGB data sacrificing information. In contrast, RAW images offer richer representation, which is crucial for precise recognition, particularly in challenging conditions…

Computer Vision and Pattern Recognition · Computer Science 2024-11-21 Christoph Reinders , Radu Berdan , Beril Besbinar , Junji Otsuka , Daisuke Iso

Recently, reference-based image super-resolution (RefSR) has shown excellent performance in image super-resolution (SR) tasks. The main idea of RefSR is to utilize additional information from the reference (Ref) image to recover the…

Computer Vision and Pattern Recognition · Computer Science 2024-01-30 Jeongho Min , Yejun Lee , Dongyoung Kim , Jaejun Yoo

Cross-domain mapping has been a very active topic in recent years. Given one image, its main purpose is to translate it to the desired target domain, or multiple domains in the case of multiple labels. This problem is highly challenging due…

Computer Vision and Pattern Recognition · Computer Science 2019-09-06 Andrés Romero , Pablo Arbeláez , Luc Van Gool , Radu Timofte

Recently, image-to-image translation research has witnessed remarkable progress. Although current approaches successfully generate diverse outputs or perform scalable image transfer, these properties have not been combined into a single…

Computer Vision and Pattern Recognition · Computer Science 2019-08-20 Yaxing Wang , Abel Gonzalez-Garcia , Joost van de Weijer , Luis Herranz

Reliable perception during fast motion maneuvers or in high dynamic range environments is crucial for robotic systems. Since event cameras are robust to these challenging conditions, they have great potential to increase the reliability of…

Computer Vision and Pattern Recognition · Computer Science 2022-02-04 Nico Messikommer , Daniel Gehrig , Mathias Gehrig , Davide Scaramuzza

Real-world robotics problems often occur in domains that differ significantly from the robot's prior training environment. For many robotic control tasks, real world experience is expensive to obtain, but data is easy to collect in either…

Computer Vision and Pattern Recognition · Computer Science 2017-05-29 Eric Tzeng , Coline Devin , Judy Hoffman , Chelsea Finn , Pieter Abbeel , Sergey Levine , Kate Saenko , Trevor Darrell

Deep learning-based automatic sleep staging has significantly advanced in performance and plays a crucial role in the diagnosis of sleep disorders. However, those models often struggle to generalize on unseen subjects due to variability in…

Machine Learning · Computer Science 2025-10-15 Sangmin Jo , Jee Seok Yoon , Wootaek Jeong , Kwanseok Oh , Heung-Il Suk

Performance on benchmark datasets has drastically improved with advances in deep learning. Still, cross-dataset generalization performance remains relatively low due to the domain shift that can occur between two different datasets. This…

Computer Vision and Pattern Recognition · Computer Science 2019-01-08 Alexandra Carlson , Katherine A. Skinner , Ram Vasudevan , Matthew Johnson-Roberson

Recent image-to-image translation models have shown great success in mapping local textures between two domains. Existing approaches rely on a cycle-consistency constraint that supervises the generators to learn an inverse mapping. However,…

Computer Vision and Pattern Recognition · Computer Science 2021-12-15 Wenju Xu , Guanghui Wang

High-quality imaging of dynamic scenes in extremely low-light conditions is highly challenging. Photon scarcity induces severe noise and texture loss, causing significant image degradation. Event cameras, featuring a high dynamic range (120…

Computer Vision and Pattern Recognition · Computer Science 2026-03-23 Haoyue Liu , Jinghan Xu , Luxin Feng , Hanyu Zhou , Haozhi Zhao , Yi Chang , Luxin Yan

Unpaired Image-to-image Translation is a new rising and challenging vision problem that aims to learn a mapping between unaligned image pairs in diverse domains. Recent advances in this field like MUNIT and DRIT mainly focus on…

Computer Vision and Pattern Recognition · Computer Science 2019-05-07 Zhiqiang Shen , Mingyang Huang , Jianping Shi , Xiangyang Xue , Thomas Huang

Comparing images captured by disparate sensors is a common challenge in remote sensing. This requires image translation -- converting imagery from one sensor domain to another while preserving the original content. Denoising Diffusion…

Computer Vision and Pattern Recognition · Computer Science 2024-12-05 João Gabriel Vinholi , Marco Chini , Anis Amziane , Renato Machado , Danilo Silva , Patrick Matgen

Multi-modal Magnetic Resonance Imaging (MRI) translation leverages information from source MRI sequences to generate target modalities, enabling comprehensive diagnosis while overcoming the limitations of acquiring all sequences. While…

Image and Video Processing · Electrical Eng. & Systems 2025-05-20 Jiyao Liu , Shangqi Gao , Yuxin Li , Lihao Liu , Xin Gao , Zhaohu Xing , Junzhi Ning , Yanzhou Su , Xiao-Yong Zhang , Junjun He , Ningsheng Xu , Xiahai Zhuang

Most vision models are trained on RGB images processed through ISP pipelines optimized for human perception, which can discard sensor-level information useful for machine reasoning. RAW images preserve unprocessed scene data, enabling…

Computer Vision and Pattern Recognition · Computer Science 2026-02-11 Mishal Fatima , Shashank Agnihotri , Kanchana Vaishnavi Gandikota , Michael Moeller , Margret Keuper

RAW images preserve superior fidelity and rich scene information compared to RGB, making them essential for tasks in challenging imaging conditions. To alleviate the high cost of data collection, recent RGB-to-RAW conversion methods aim to…

Computer Vision and Pattern Recognition · Computer Science 2026-03-17 Huanjing Yue , Shangbin Xie , Cong Cao , Qian Wu , Lei Zhang , Lei Zhao , Jingyu Yang

Image translation across different domains has attracted much attention in both machine learning and computer vision communities. Taking the translation from source domain $\mathcal{D}_s$ to target domain $\mathcal{D}_t$ as an example,…

Computer Vision and Pattern Recognition · Computer Science 2019-05-30 Jianxin Lin , Yingce Xia , Yijun Wang , Tao Qin , Zhibo Chen

Neural Radiance Fields (NeRF) use multi-view images for 3D scene representation, demonstrating remarkable performance. As one of the primary sources of multi-view images, multi-camera systems encounter challenges such as varying intrinsic…

Computer Vision and Pattern Recognition · Computer Science 2024-12-09 Yu Gao , Lutong Su , Hao Liang , Yufeng Yue , Yi Yang , Mengyin Fu

Self-supervised pretrain techniques have been widely used to improve the downstream tasks' performance. However, real-world magnetic resonance (MR) studies usually consist of different sets of contrasts due to different acquisition…

Image and Video Processing · Electrical Eng. & Systems 2025-06-17 Badhan Kumar Das , Ajay Singh , Gengyan Zhao , Han Liu , Thomas J. Re , Dorin Comaniciu , Eli Gibson , Andreas Maier

Real-world image super-resolution (Real SR) aims to generate high-fidelity, detail-rich high-resolution (HR) images from low-resolution (LR) counterparts. Existing Real SR methods primarily focus on generating details from the LR RGB…

Image and Video Processing · Electrical Eng. & Systems 2024-11-22 Long Peng , Wenbo Li , Jiaming Guo , Xin Di , Haoze Sun , Yong Li , Renjing Pei , Yang Wang , Yang Cao , Zheng-Jun Zha

Diffusion models achieved great success in image synthesis, but still face challenges in high-resolution generation. Through the lens of discrete cosine transformation, we find the main reason is that \emph{the same noise level on a higher…

Computer Vision and Pattern Recognition · Computer Science 2023-09-08 Jiayan Teng , Wendi Zheng , Ming Ding , Wenyi Hong , Jianqiao Wangni , Zhuoyi Yang , Jie Tang