English
Related papers

Related papers: Occupancy-map-based rate distortion optimization f…

200 papers

3D perception in point clouds is transforming the perception ability of future intelligent machines. Point cloud algorithms, however, are plagued by irregular memory accesses, leading to massive inefficiencies in the memory sub-system,…

Hardware Architecture · Computer Science 2022-04-25 Yu Feng , Gunnar Hammonds , Yiming Gan , Yuhao Zhu

Sparse representation leads to an efficient way to approximately recover a signal by the linear composition of a few bases from a learnt dictionary, based on which various successful applications have been achieved. However, in the scenario…

Computer Vision and Pattern Recognition · Computer Science 2018-05-04 Xiang Zhang , Jiarui Sun , Siwei Ma , Zhouchen Lin , Jian Zhang , Shiqi Wang , Wen Gao

Representing scenes from multi-view images is a crucial task in computer vision with extensive applications. However, inherent photometric distortions in the camera imaging can significantly degrade image quality. Without accounting for…

Computer Vision and Pattern Recognition · Computer Science 2025-06-27 Weichen Dai , Kangcheng Ma , Jiaxin Wang , Kecen Pan , Yuhang Ming , Hua Zhang , Wanzeng Kong

We consider the attributes of a point cloud as samples of a vector-valued volumetric function at discrete positions. To compress the attributes given the positions, we compress the parameters of the volumetric function. We model the…

Graphics · Computer Science 2021-11-18 Berivan Isik , Philip A. Chou , Sung Jin Hwang , Nick Johnston , George Toderici

We propose a method to compress full-resolution video sequences with implicit neural representations. Each frame is represented as a neural network that maps coordinate positions to pixel values. We use a separate implicit network to…

Machine Learning · Computer Science 2021-12-22 Yunfan Zhang , Ties van Rozendaal , Johann Brehmer , Markus Nagel , Taco Cohen

Holographic representations of data enable distributed storage with progressive refinement when the stored packets of data are made available in any arbitrary order. In this paper, we propose and test patch-based transform coding…

Information Theory · Computer Science 2021-03-01 Alfred Marcel Bruckstein , Martianus Frederic Ezerman , Adamas Aqsa Fahreza , San Ling

There are many tasks within video compression which require fast bit rate estimation. As an example, rate-control algorithms are only feasible because it is possible to estimate the required bit rate without needing to encode the entire…

Image and Video Processing · Electrical Eng. & Systems 2022-02-16 Fabian Brand , Christian Herglotz , André Kaup

Inpainting-based image compression is a promising alternative to classical transform-based lossy codecs. Typically it stores a carefully selected subset of all pixel locations and their colour values. In the decoding phase the missing…

Image and Video Processing · Electrical Eng. & Systems 2023-05-16 Ferdinand Jost , Vassillen Chizhov , Joachim Weickert

Depth maps are needed by various graphics rendering and processing operations. Depth map streaming is often necessary when such operations are performed in a distributed system and it requires in most cases fast performing compression,…

Multimedia · Computer Science 2022-07-01 Matti Siekkinen , Teemu Kämäräinen

3D Gaussian Splatting (3DGS) is a powerful reconstruction technique, but it needs to be initialized from accurate camera poses and high-fidelity point clouds. Typically, the initialization is taken from Structure-from-Motion (SfM)…

Computer Vision and Pattern Recognition · Computer Science 2025-04-15 Jizong Peng , Tze Ho Elden Tse , Kai Xu , Wenchao Gao , Angela Yao

Autonomous vehicles rely on LiDAR sensors to generate 3D point clouds for accurate segmentation and object detection. In a context of a smart city framework, we would like to understand the effect that transmission (compression) can have on…

Image and Video Processing · Electrical Eng. & Systems 2025-09-30 Tiago de S. Fernandes , Ricardo L. de Queiroz

Recently, 3D Gaussian Spatting (3DGS) has gained widespread attention in Novel View Synthesis (NVS) due to the remarkable real-time rendering performance. However, the substantial cost of storage and transmission of vanilla 3DGS hinders its…

Computer Vision and Pattern Recognition · Computer Science 2025-04-01 Jingui Ma , Yang Hu , Luyang Tang , Jiayu Yang , Yongqi Zhai , Ronggang Wang

We introduce a new approach for generating realistic 3D models with UV maps through a representation termed "Object Images." This approach encapsulates surface geometry, appearance, and patch structures within a 64x64 pixel image,…

Computer Vision and Pattern Recognition · Computer Science 2024-08-07 Xingguang Yan , Han-Hung Lee , Ziyu Wan , Angel X. Chang

Point cloud is a fundamental 3D representation which is widely used in real world applications such as autonomous driving. As a newly-developed media format which is characterized by complexity and irregularity, point cloud creates a need…

Computer Vision and Pattern Recognition · Computer Science 2019-05-10 Wei Yan , Yiting shao , Shan Liu , Thomas H Li , Zhu Li , Ge Li

Deep neural network models used for medical image segmentation are large because they are trained with high-resolution three-dimensional (3D) images. Graphics processing units (GPUs) are widely used to accelerate the trainings. However, the…

Machine Learning · Computer Science 2018-12-20 Haruki Imai , Samuel Matzek , Tung D. Le , Yasushi Negishi , Kiyokuni Kawachiya

Neural representations for 3D meshes are emerging as an effective solution for compact storage and efficient processing. Existing methods often rely on neural overfitting, where a coarse mesh is stored and progressively refined through…

Computer Vision and Pattern Recognition · Computer Science 2025-11-25 Xiang Gao , Yuanpeng Liu , Xinmu Wang , Jiazhi Li , Minghao Guo , Yu Guo , Xiyun Song , Heather Yu , Zhiqiang Lao , Xianfeng David Gu

Modern video codecs and learning-based approaches struggle for semantic reconstruction at extremely low bit-rates due to reliance on low-level spatiotemporal redundancies. Generative models, especially diffusion models, offer a new paradigm…

Image and Video Processing · Electrical Eng. & Systems 2026-02-06 Maojun Zhang , Haotian Wu , Richeng Jin , Deniz Gunduz , Krystian Mikolajczyk

Deep learning is increasingly being used to perform machine vision tasks such as classification, object detection, and segmentation on 3D point cloud data. However, deep learning inference is computationally expensive. The limited…

Image and Video Processing · Electrical Eng. & Systems 2023-08-14 Mateen Ulhaq , Ivan V. Bajić

Neural implicit surfaces have become an important technique for multi-view 3D reconstruction but their accuracy remains limited. In this paper, we argue that this comes from the difficulty to learn and render high frequency textures with…

Computer Vision and Pattern Recognition · Computer Science 2022-05-10 François Darmon , Bénédicte Bascle , Jean-Clément Devaux , Pascal Monasse , Mathieu Aubry

This paper describes a novel lossless point cloud compression algorithm that uses a neural network for estimating the coding probabilities for the occupancy status of voxels, depending on wide three dimensional contexts around the voxel to…

Computer Vision and Pattern Recognition · Computer Science 2021-10-12 Emre Can Kaya , Ioan Tabus