中文
相关论文

相关论文: Pix2Next: Leveraging Vision Foundation Models for …

200 篇论文

In the field of computer vision, visible light images often exhibit low contrast in low-light conditions, presenting a significant challenge. While infrared imagery provides a potential solution, its utilization entails high costs and…

计算机视觉与模式识别 · 计算机科学 2024-04-30 Yijia Chen , Pinghua Chen , Xiangxin Zhou , Yingtie Lei , Ziyang Zhou , Mingxian Li

Generative Adversarial Networks (GANs) have significantly advanced image processing, with Pix2Pix being a notable framework for image-to-image translation. This paper explores a novel application of Pix2Pix to transform abstract map images…

计算机视觉与模式识别 · 计算机科学 2024-05-02 Zhenglin Li , Bo Guan , Yuanzhou Wei , Yiming Zhou , Jingyu Zhang , Jinxin Xu

Recent progress in computational photography has shown that we can acquire near-infrared (NIR) information in addition to the normal visible (RGB) band, with only slight modifications to standard digital cameras. Due to the proximity of the…

计算机视觉与模式识别 · 计算机科学 2014-06-25 Neda Salamati , Diane Larlus , Gabriela Csurka , Sabine Süsstrunk

Semantic analysis on visible (RGB) and infrared (IR) images has gained significant attention due to their enhanced accuracy and robustness under challenging conditions including low-illumination and adverse weather. However, due to the lack…

计算机视觉与模式识别 · 计算机科学 2025-10-14 Maoxun Yuan , Bo Cui , Tianyi Zhao , Jiayi Wang , Shan Fu , Xue Yang , Xingxing Wei

We present Pix2Gif, a motion-guided diffusion model for image-to-GIF (video) generation. We tackle this problem differently by formulating the task as an image translation problem steered by text and motion magnitude prompts, as shown in…

计算机视觉与模式识别 · 计算机科学 2024-03-11 Hitesh Kandala , Jianfeng Gao , Jianwei Yang

Seeing-in-the-dark is one of the most important and challenging computer vision tasks due to its wide applications and extreme complexities of in-the-wild scenarios. Existing arts can be mainly divided into two threads: 1) RGB-dependent…

计算机视觉与模式识别 · 计算机科学 2023-03-22 Muyao Niu , Zhuoxiao Li , Zhihang Zhong , Yinqiang Zheng

This study aims to learn a translation from visible to infrared imagery, bridging the domain gap between the two modalities so as to improve accuracy on downstream tasks including object detection. Previous approaches attempt to perform…

计算机视觉与模式识别 · 计算机科学 2024-08-06 Prahlad Anand , Qiranul Saadiyean , Aniruddh Sikdar , Nalini N , Suresh Sundaram

Current deep learning approaches in computer vision primarily focus on RGB data sacrificing information. In contrast, RAW images offer richer representation, which is crucial for precise recognition, particularly in challenging conditions…

计算机视觉与模式识别 · 计算机科学 2024-11-21 Christoph Reinders , Radu Berdan , Beril Besbinar , Junji Otsuka , Daisuke Iso

The current industry practice for 24-hour outdoor imaging is to use a silicon camera supplemented with near-infrared (NIR) illumination. This will result in color images with poor contrast at daytime and absence of chrominance at nighttime.…

图像与视频处理 · 电气工程与系统科学 2020-05-12 Feifan Lv , Yinqiang Zheng , Yicheng Li , Feng Lu

The diffusion model has demonstrated superior performance in synthesizing diverse and high-quality images for text-guided image translation. However, there remains room for improvement in both the formulation of text prompts and the…

计算机视觉与模式识别 · 计算机科学 2025-03-27 Qi Si , Bo Wang , Zhao Zhang

Large-scale text-to-image generative models have shown their remarkable ability to synthesize diverse and high-quality images. However, it is still challenging to directly apply these models for editing real images for two reasons. First,…

计算机视觉与模式识别 · 计算机科学 2023-02-07 Gaurav Parmar , Krishna Kumar Singh , Richard Zhang , Yijun Li , Jingwan Lu , Jun-Yan Zhu

We propose a framework for aligning and fusing multiple images into a single view using neural image representations (NIRs), also known as implicit or coordinate-based neural representations. Our framework targets burst images that exhibit…

计算机视觉与模式识别 · 计算机科学 2022-07-22 Seonghyeon Nam , Marcus A. Brubaker , Michael S. Brown

Image-to-image translation (I2I) aims at transferring the content representation from an input domain to an output one, bouncing along different target domains. Recent I2I generative models, which gain outstanding results in this task,…

计算机视觉与模式识别 · 计算机科学 2022-10-11 Eleonora Grassucci , Luigi Sigillo , Aurelio Uncini , Danilo Comminiello

Existing frameworks for image stitching often provide visually reasonable stitchings. However, they suffer from blurry artifacts and disparities in illumination, depth level, etc. Although the recent learning-based stitchings relax such…

计算机视觉与模式识别 · 计算机科学 2024-01-23 Minsu Kim , Jaewon Lee , Byeonghun Lee , Sunghoon Im , Kyong Hwan Jin

Image enhancement using the visible (V) and near-infrared (NIR) usually enhances useful image details. The enhanced images are evaluated by observers perception, instead of quantitative feature evaluation. Thus, can we say that these…

计算机视觉与模式识别 · 计算机科学 2016-08-25 Vivek Sharma , Luc Van Gool

With the strong robusticity on illumination variations, near-infrared (NIR) can be an effective and essential complement to visible (VIS) facial expression recognition in low lighting or complete darkness conditions. However, facial…

计算机视觉与模式识别 · 计算机科学 2023-12-12 Bingjun Luo , Haowen Wang , Jinpeng Wang , Junjie Zhu , Xibin Zhao , Yue Gao

Image-to-image translation (I2I), and particularly its subfield of appearance transfer, which seeks to alter the visual appearance between images while maintaining structural coherence, presents formidable challenges. Despite significant…

计算机视觉与模式识别 · 计算机科学 2023-11-29 Yuteng Ye , Guanwen Li , Hang Zhou , Cai Jiale , Junqing Yu , Yawei Luo , Zikai Song , Qilong Xing , Youjia Zhang , Wei Yang

We present HAQAGen, a unified generative model for resolution-invariant NIR-to-RGB colorization that balances chromatic realism with structural fidelity. The proposed model introduces (i) a combined loss term aligning the global color…

计算机视觉与模式识别 · 计算机科学 2026-01-06 Abhinav Attri , Rajeev Ranjan Dwivedi , Samiran Das , Vinod Kumar Kurmi

RGB-NIR fusion is a promising method for low-light imaging. However, high-intensity noise in low-light images amplifies the effect of structure inconsistency between RGB-NIR images, which fails existing algorithms. To handle this, we…

计算机视觉与模式识别 · 计算机科学 2023-04-20 Shuangping Jin , Bingbing Yu , Minhao Jing , Yi Zhou , Jiajun Liang , Renhe Ji

The long-coveted task of reconstructing 3D geometry from images is still a standing problem. In this paper, we build on the power of neural networks and introduce Pix2Vex, a network trained to convert camera-captured images into 3D…

计算机视觉与模式识别 · 计算机科学 2019-05-28 Felix Petersen , Amit H. Bermano , Oliver Deussen , Daniel Cohen-Or