中文
相关论文

相关论文: Light Field Compression with Disparity Guided Spar…

200 篇论文

Feature detectors and descriptors are key low-level vision tools that many higher-level tasks build on. Unfortunately these fail in the presence of challenging light transport effects including partial occlusion, low contrast, and…

计算机视觉与模式识别 · 计算机科学 2019-01-15 Donald G. Dansereau , Bernd Girod , Gordon Wetzstein

Light field images capture multi-view scene information and play a crucial role in 3D scene reconstruction. However, their high-dimensional nature results in enormous data volumes, posing a significant challenge for efficient compression in…

图像与视频处理 · 电气工程与系统科学 2025-10-20 Gai Zhang , Xinfeng Zhang , Lv Tang , Hongyu An , Li Zhang , Qingming Huang

Sparse decomposition has been widely used for different applications, such as source separation, image classification, image denoising and more. This paper presents a new algorithm for segmentation of an image into background and foreground…

计算机视觉与模式识别 · 计算机科学 2016-07-28 Shervin Minaee , Yao Wang

The rise of new video modalities like virtual reality or autonomous driving has increased the demand for efficient multi-view video compression methods, both in terms of rate-distortion (R-D) performance and in terms of delay and runtime.…

计算机视觉与模式识别 · 计算机科学 2024-03-27 Qiqi Hou , Farzad Farhadzadeh , Amir Said , Guillaume Sautiere , Hoang Le

We study causal, low-latency, sequential video compression when the output is subjected to both a mean squared-error (MSE) distortion loss as well as a perception loss to target realism. Motivated by prior approaches, we consider two…

图像与视频处理 · 电气工程与系统科学 2023-08-24 Sadaf Salehkalaibar , Buu Phan , Jun Chen , Wei Yu , Ashish Khisti

Autoencoder-based image codecs achieve state-of-the-art compression performance but often incur high computational complexity, particularly at decoding time. This work introduces a low-complexity learned image compression framework based on…

图像与视频处理 · 电气工程与系统科学 2026-05-14 Théophile Blard , Pierrick Philippe , Théo Ladune , Xiaoran Jiang , Olivier Déforges

Convolutional Sparse Coding (CSC) is a well-established image representation model especially suited for image restoration tasks. In this work, we extend the applicability of this model by proposing a supervised approach to convolutional…

计算机视觉与模式识别 · 计算机科学 2018-04-10 Lama Affara , Bernard Ghanem , Peter Wonka

The high-dimensional nature of the 4-D light field (LF) poses great challenges in achieving efficient and effective feature embedding, that severely impacts the performance of downstream tasks. To tackle this crucial issue, in contrast to…

图像与视频处理 · 电气工程与系统科学 2024-01-12 Xianqiang Lyu , Junhui Hou

Sparse representations of images are useful in many computer vision applications. Sparse coding with an $l_1$ penalty and a learned linear dictionary requires regularization of the dictionary to prevent a collapse in the $l_1$ norms of the…

计算机视觉与模式识别 · 计算机科学 2022-09-09 Katrina Evtimova , Yann LeCun

In this paper, we delve into the realm of 4-D light fields (LFs) to enhance underwater imaging plagued by light absorption, scattering, and other challenges. Contrasting with conventional 2-D RGB imaging, 4-D LF imaging excels in capturing…

计算机视觉与模式识别 · 计算机科学 2025-07-15 Yuji Lin , Junhui Hou , Xianqiang Lyu , Qian Zhao , Deyu Meng

In the Fourier domain, luminance information is primarily encoded in the amplitude spectrum, while spatial structures are captured in the phase components. The traditional Fourier Frequency information fitting employs pixel-wise loss…

计算机视觉与模式识别 · 计算机科学 2025-10-03 Yan Xingyang , Huang Xiaohong , Zhang Zhao , You Tian , Xu Ziheng

Light fields are 4D scene representation typically structured as arrays of views, or several directional samples per pixel in a single view. This highly correlated structure is not very efficient to transmit and manipulate (especially for…

计算机视觉与模式识别 · 计算机科学 2021-03-23 Menghan Xia , Jose Echevarria , Minshan Xie , Tien-Tsin Wong

Accurate 3D reconstruction of colored objects with structured light (SL) is hindered by lateral chromatic aberration (LCA) in optical components and uneven noise characteristics across RGB channels. This paper introduces lateral chromatic…

计算机视觉与模式识别 · 计算机科学 2026-03-12 Wonbeen Oh , Jae-Sang Hyun

We present a novel feature selection technique, Sparse Linear Centroid-Encoder (SLCE). The algorithm uses a linear transformation to reconstruct a point as its class centroid and, at the same time, uses the $\ell_1$-norm penalty to filter…

机器学习 · 计算机科学 2023-06-12 Tomojit Ghosh , Michael Kirby , Karim Karimov

Plenoptic cameras offer a cost effective solution to capture light fields by multiplexing multiple views on a single image sensor. However, the high angular resolution is achieved at the expense of reducing the spatial resolution of each…

计算机视觉与模式识别 · 计算机科学 2018-09-28 Reuben A. Farrugia , C. Guillemot

Light field (LF) cameras record both intensity and directions of light rays, and capture scenes from a number of viewpoints. Both information within each perspective (i.e., spatial information) and among different perspectives (i.e.,…

图像与视频处理 · 电气工程与系统科学 2020-06-05 Yingqian Wang , Longguang Wang , Jungang Yang , Wei An , Jingyi Yu , Yulan Guo

Dual pixels contain disparity cues arising from the defocus blur. This disparity information is useful for many vision tasks ranging from autonomous driving to 3D creative realism. However, directly estimating disparity from dual pixels is…

计算机视觉与模式识别 · 计算机科学 2024-05-21 Aryan Garg , Raghav Mallampali , Akshat Joshi , Shrisudhan Govindarajan , Kaushik Mitra

Long video understanding is a complex task that requires both spatial detail and temporal awareness. While Vision-Language Models (VLMs) obtain frame-level understanding capabilities through multi-frame input, they suffer from information…

计算机视觉与模式识别 · 计算机科学 2025-04-10 Ziyi Wang , Haoran Wu , Yiming Rong , Deyang Jiang , Yixin Zhang , Yunlong Zhao , Shuang Xu , Bo XU

To enrich the functionalities of traditional cameras, light field cameras record both the intensity and direction of light rays, so that images can be rendered with user-defined camera parameters via computations. The added capability and…

图像与视频处理 · 电气工程与系统科学 2022-01-25 Muhammad Umair Mukati , Xi Zhang , Xiaolin Wu , Søren Forchhammer

Vision Transformer (ViT)-based sparse multi-view 3D object detectors have achieved remarkable accuracy but still suffer from high inference latency due to heavy token processing. To accelerate these models, token compression has been widely…

计算机视觉与模式识别 · 计算机科学 2026-04-17 Mingqian Ji , Shanshan Zhang , Jian Yang