English
Related papers

Related papers: Light Field Compression with Disparity Guided Spar…

200 papers

Feature detectors and descriptors are key low-level vision tools that many higher-level tasks build on. Unfortunately these fail in the presence of challenging light transport effects including partial occlusion, low contrast, and…

Computer Vision and Pattern Recognition · Computer Science 2019-01-15 Donald G. Dansereau , Bernd Girod , Gordon Wetzstein

Light field images capture multi-view scene information and play a crucial role in 3D scene reconstruction. However, their high-dimensional nature results in enormous data volumes, posing a significant challenge for efficient compression in…

Image and Video Processing · Electrical Eng. & Systems 2025-10-20 Gai Zhang , Xinfeng Zhang , Lv Tang , Hongyu An , Li Zhang , Qingming Huang

Sparse decomposition has been widely used for different applications, such as source separation, image classification, image denoising and more. This paper presents a new algorithm for segmentation of an image into background and foreground…

Computer Vision and Pattern Recognition · Computer Science 2016-07-28 Shervin Minaee , Yao Wang

The rise of new video modalities like virtual reality or autonomous driving has increased the demand for efficient multi-view video compression methods, both in terms of rate-distortion (R-D) performance and in terms of delay and runtime.…

Computer Vision and Pattern Recognition · Computer Science 2024-03-27 Qiqi Hou , Farzad Farhadzadeh , Amir Said , Guillaume Sautiere , Hoang Le

We study causal, low-latency, sequential video compression when the output is subjected to both a mean squared-error (MSE) distortion loss as well as a perception loss to target realism. Motivated by prior approaches, we consider two…

Image and Video Processing · Electrical Eng. & Systems 2023-08-24 Sadaf Salehkalaibar , Buu Phan , Jun Chen , Wei Yu , Ashish Khisti

Autoencoder-based image codecs achieve state-of-the-art compression performance but often incur high computational complexity, particularly at decoding time. This work introduces a low-complexity learned image compression framework based on…

Image and Video Processing · Electrical Eng. & Systems 2026-05-14 Théophile Blard , Pierrick Philippe , Théo Ladune , Xiaoran Jiang , Olivier Déforges

Convolutional Sparse Coding (CSC) is a well-established image representation model especially suited for image restoration tasks. In this work, we extend the applicability of this model by proposing a supervised approach to convolutional…

Computer Vision and Pattern Recognition · Computer Science 2018-04-10 Lama Affara , Bernard Ghanem , Peter Wonka

The high-dimensional nature of the 4-D light field (LF) poses great challenges in achieving efficient and effective feature embedding, that severely impacts the performance of downstream tasks. To tackle this crucial issue, in contrast to…

Image and Video Processing · Electrical Eng. & Systems 2024-01-12 Xianqiang Lyu , Junhui Hou

Sparse representations of images are useful in many computer vision applications. Sparse coding with an $l_1$ penalty and a learned linear dictionary requires regularization of the dictionary to prevent a collapse in the $l_1$ norms of the…

Computer Vision and Pattern Recognition · Computer Science 2022-09-09 Katrina Evtimova , Yann LeCun

In this paper, we delve into the realm of 4-D light fields (LFs) to enhance underwater imaging plagued by light absorption, scattering, and other challenges. Contrasting with conventional 2-D RGB imaging, 4-D LF imaging excels in capturing…

Computer Vision and Pattern Recognition · Computer Science 2025-07-15 Yuji Lin , Junhui Hou , Xianqiang Lyu , Qian Zhao , Deyu Meng

In the Fourier domain, luminance information is primarily encoded in the amplitude spectrum, while spatial structures are captured in the phase components. The traditional Fourier Frequency information fitting employs pixel-wise loss…

Computer Vision and Pattern Recognition · Computer Science 2025-10-03 Yan Xingyang , Huang Xiaohong , Zhang Zhao , You Tian , Xu Ziheng

Light fields are 4D scene representation typically structured as arrays of views, or several directional samples per pixel in a single view. This highly correlated structure is not very efficient to transmit and manipulate (especially for…

Computer Vision and Pattern Recognition · Computer Science 2021-03-23 Menghan Xia , Jose Echevarria , Minshan Xie , Tien-Tsin Wong

Accurate 3D reconstruction of colored objects with structured light (SL) is hindered by lateral chromatic aberration (LCA) in optical components and uneven noise characteristics across RGB channels. This paper introduces lateral chromatic…

Computer Vision and Pattern Recognition · Computer Science 2026-03-12 Wonbeen Oh , Jae-Sang Hyun

We present a novel feature selection technique, Sparse Linear Centroid-Encoder (SLCE). The algorithm uses a linear transformation to reconstruct a point as its class centroid and, at the same time, uses the $\ell_1$-norm penalty to filter…

Machine Learning · Computer Science 2023-06-12 Tomojit Ghosh , Michael Kirby , Karim Karimov

Plenoptic cameras offer a cost effective solution to capture light fields by multiplexing multiple views on a single image sensor. However, the high angular resolution is achieved at the expense of reducing the spatial resolution of each…

Computer Vision and Pattern Recognition · Computer Science 2018-09-28 Reuben A. Farrugia , C. Guillemot

Light field (LF) cameras record both intensity and directions of light rays, and capture scenes from a number of viewpoints. Both information within each perspective (i.e., spatial information) and among different perspectives (i.e.,…

Image and Video Processing · Electrical Eng. & Systems 2020-06-05 Yingqian Wang , Longguang Wang , Jungang Yang , Wei An , Jingyi Yu , Yulan Guo

Dual pixels contain disparity cues arising from the defocus blur. This disparity information is useful for many vision tasks ranging from autonomous driving to 3D creative realism. However, directly estimating disparity from dual pixels is…

Computer Vision and Pattern Recognition · Computer Science 2024-05-21 Aryan Garg , Raghav Mallampali , Akshat Joshi , Shrisudhan Govindarajan , Kaushik Mitra

Long video understanding is a complex task that requires both spatial detail and temporal awareness. While Vision-Language Models (VLMs) obtain frame-level understanding capabilities through multi-frame input, they suffer from information…

Computer Vision and Pattern Recognition · Computer Science 2025-04-10 Ziyi Wang , Haoran Wu , Yiming Rong , Deyang Jiang , Yixin Zhang , Yunlong Zhao , Shuang Xu , Bo XU

To enrich the functionalities of traditional cameras, light field cameras record both the intensity and direction of light rays, so that images can be rendered with user-defined camera parameters via computations. The added capability and…

Image and Video Processing · Electrical Eng. & Systems 2022-01-25 Muhammad Umair Mukati , Xi Zhang , Xiaolin Wu , Søren Forchhammer

Vision Transformer (ViT)-based sparse multi-view 3D object detectors have achieved remarkable accuracy but still suffer from high inference latency due to heavy token processing. To accelerate these models, token compression has been widely…

Computer Vision and Pattern Recognition · Computer Science 2026-04-17 Mingqian Ji , Shanshan Zhang , Jian Yang