English
Related papers

Related papers: Real-Time Computer-Generated EIA for Light Field D…

200 papers

Replicating In-Context Learning (ICL) in computer vision remains challenging due to task heterogeneity. We propose \textbf{VIRAL}, a framework that elicits visual reasoning from a pre-trained image editing model by formulating ICL as…

Computer Vision and Pattern Recognition · Computer Science 2026-02-04 Zhiwen Li , Zhongjie Duan , Jinyan Ye , Cen Chen , Daoyuan Chen , Yaliang Li , Yingda Chen

This paper presents a learning-based approach to synthesize the view from an arbitrary camera position given a sparse set of images. A key challenge for this novel view synthesis arises from the reconstruction process, when the views from…

Computer Vision and Pattern Recognition · Computer Science 2021-04-06 Nan Meng , Kai Li , Jianzhuang Liu , Edmund Y. Lam

Light field cameras have many advantages over traditional cameras, as they allow the user to change various camera settings after capture. However, capturing light fields requires a huge bandwidth to record the data: a modern light field…

Computer Vision and Pattern Recognition · Computer Science 2017-05-09 Ting-Chun Wang , Jun-Yan Zhu , Nima Khademi Kalantari , Alexei A. Efros , Ravi Ramamoorthi

Distinguishing between real and AI-generated images, commonly referred to as 'image detection', presents a timely and significant challenge. Despite extensive research in the (semi-)supervised regime, zero-shot and few-shot solutions have…

Computer Vision and Pattern Recognition · Computer Science 2025-04-23 Jonathan Brokman , Amit Giloni , Omer Hofman , Roman Vainshtein , Hisashi Kojima , Guy Gilboa

Most of the artificial lights fluctuate in response to the grid's alternating current and exhibit subtle variations in terms of both intensity and spectrum, providing the potential to estimate the Electric Network Frequency (ENF) from…

Image and Video Processing · Electrical Eng. & Systems 2023-05-05 Lexuan Xu , Guang Hua , Haijian Zhang , Lei Yu , Ning Qiao

Current methods for multimodal medical imaging based disease recognition face two major challenges. First, the prevailing "fusion after unimodal image embedding" paradigm cannot fully leverage the complementary and correlated information in…

Computer Vision and Pattern Recognition · Computer Science 2026-04-07 Qijie Wei , Hailan Lin , Xirong Li

Over the past few years, Text-to-Image (T2I) generation approaches based on diffusion models have gained significant attention. However, vanilla diffusion models often suffer from spelling inaccuracies in the text displayed within the…

Computer Vision and Pattern Recognition · Computer Science 2024-10-30 Sanyam Lakhanpal , Shivang Chopra , Vinija Jain , Aman Chadha , Man Luo

There is a general expectation that robots should operate in urban environments often consisting of potentially dynamic entities including people, furniture and automobiles. Dynamic objects pose challenges to visual SLAM algorithms by…

Robotics · Computer Science 2020-12-22 Pushyami Kaveti , Hanumant Singh

Deep learning has bolstered gaze estimation techniques, but real-world deployment has been impeded by inadequate training datasets. This problem is exacerbated by both hardware-induced variations in eye images and inherent biological…

Computer Vision and Pattern Recognition · Computer Science 2025-05-01 Sean Anthony Byrne , Virmarie Maquiling , Marcus Nyström , Enkelejda Kasneci , Diederick C. Niehorster

The Visual Turing Test is the ultimate goal to evaluate the realism of holographic displays. Previous studies have focused on addressing challenges such as limited \'etendue and image quality over a large focal volume, but they have not…

Computer Vision and Pattern Recognition · Computer Science 2023-07-13 Florian Schiffers , Praneeth Chakravarthula , Nathan Matsuda , Grace Kuo , Ethan Tseng , Douglas Lanman , Felix Heide , Oliver Cossairt

In recent years, many design automation methods have been developed to routinely create approximate implementations of circuits and programs that show excellent trade-offs between the quality of output and required resources. This paper…

Neural and Evolutionary Computing · Computer Science 2021-08-17 Lukas Sekanina

Ensemble learning is a method of combining multiple trained models to improve model accuracy. We propose the usage of such methods, specifically ensemble average, inside Convolutional Neural Network (CNN) architectures by replacing the…

Machine Learning · Computer Science 2019-08-08 Abduallah Mohamed , Xinrui Hua , Xianda Zhou , Christian Claudel

Image aesthetic assessment (IAA) aims to predict the aesthetic quality of images as perceived by humans. While recent IAA models achieve strong predictive performance, they offer little insight into the factors driving their predictions.…

Computer Vision and Pattern Recognition · Computer Science 2026-03-20 Xiao-Chang Liu , Johan Wagemans

Uneven light image enhancement is a highly demanded task in many industrial image processing applications. Many existing enhancement methods using physical lighting models or deep-learning techniques often lead to unnatural results. This is…

Image and Video Processing · Electrical Eng. & Systems 2023-05-26 Tian Pu , Shuhang Wang , Zhenming Peng , Qingsong Zhu

Event camera has recently received much attention for low-light image enhancement (LIE) thanks to their distinct advantages, such as high dynamic range. However, current research is prohibitively restricted by the lack of large-scale,…

Computer Vision and Pattern Recognition · Computer Science 2024-04-02 Guoqiang Liang , Kanghao Chen , Hangyu Li , Yunfan Lu , Lin Wang

Video Variational Autoencoder (VAE) enables latent video generative modeling by mapping the visual world into compact spatiotemporal latent spaces, improving training efficiency and stability. While existing video VAEs achieve commendable…

Computer Vision and Pattern Recognition · Computer Science 2026-05-05 Yian Zhao , Feng Wang , Qiushan Guo , Chang Liu , Xiangyang Ji , Jian Zhang , Jie Chen

Computer vision is hard because of a large variability in lighting, shape, and texture; in addition the image signal is non-additive due to occlusion. Generative models promised to account for this variability by accurately modelling the…

Computer Vision and Pattern Recognition · Computer Science 2015-03-10 Varun Jampani , Sebastian Nowozin , Matthew Loper , Peter V. Gehler

Image tokenization plays a critical role in reducing the computational demands of modeling high-resolution images, significantly improving the efficiency of image and multimodal understanding and generation. Recent advances in 1D latent…

Computer Vision and Pattern Recognition · Computer Science 2025-06-27 Ze Wang , Hao Chen , Benran Hu , Jiang Liu , Ximeng Sun , Jialian Wu , Yusheng Su , Xiaodong Yu , Emad Barsoum , Zicheng Liu

This work introduces ILIAS, a new test dataset for Instance-Level Image retrieval At Scale. It is designed to evaluate the ability of current and future foundation models and retrieval techniques to recognize particular objects. The key…

In recent years, the task of video prediction-forecasting future video given past video frames-has attracted attention in the research community. In this paper we propose a novel approach to this problem with Vector Quantized Variational…

Computer Vision and Pattern Recognition · Computer Science 2021-03-03 Jacob Walker , Ali Razavi , Aäron van den Oord