English
Related papers

Related papers: OSLO-IC: On-the-Sphere Learned Omnidirectional Ima…

200 papers

Omnidirectional images are increasingly used in robotics and vision due to their wide field of view. However, extending 3D Gaussian Splatting (3DGS) to panoramic camera models remains challenging, as existing formulations are designed for…

Computer Vision and Pattern Recognition · Computer Science 2026-03-10 Zhe Yang , Guoqiang Zhao , Sheng Wu , Kai Luo , Kailun Yang

Recent years, learned image compression has made tremendous progress to achieve impressive coding efficiency. Its coding gain mainly comes from non-linear neural network-based transform and learnable entropy modeling. However, most studies…

Computer Vision and Pattern Recognition · Computer Science 2025-03-25 Donghui Feng , Zhengxue Cheng , Shen Wang , Ronghua Wu , Hongwei Hu , Guo Lu , Li Song

In recent research, Learned Image Compression has gained prominence for its capacity to outperform traditional handcrafted pipelines, especially at low bit-rates. While existing methods incorporate convolutional priors with occasional…

Image and Video Processing · Electrical Eng. & Systems 2023-10-18 Natacha Luka , Romain Negrel , David Picard

This work addresses two major issues of end-to-end learned image compression (LIC) based on deep neural networks: variable-rate learning where separate networks are required to generate compressed images with varying qualities, and the…

Image and Video Processing · Electrical Eng. & Systems 2024-04-10 Wei Jiang , Wei Wang , Songnan Li , Shan Liu

In recent years, the research community has shown a lot of interest to panoramic images that offer a 360-degree directional perspective. Multiple data modalities can be fed, and complimentary characteristics can be utilized for more robust…

Computer Vision and Pattern Recognition · Computer Science 2023-08-21 Suresh Guttikonda , Jason Rambach

Multimodal learning has gained much success in recent years. However, current multimodal fusion methods adopt the attention mechanism of Transformers to implicitly learn the underlying correlation of multimodal features. As a result, the…

Computer Vision and Pattern Recognition · Computer Science 2025-11-27 Thanh-Dat Truong , Christophe Bobda , Nitin Agarwal , Khoa Luu

Learning based video compression attracts increasing attention in the past few years. The previous hybrid coding approaches rely on pixel space operations to reduce spatial and temporal redundancy, which may suffer from inaccurate motion…

Image and Video Processing · Electrical Eng. & Systems 2021-08-24 Zhihao Hu , Guo Lu , Dong Xu

Omnidirectional image super-resolution (ODISR) aims to upscale low-resolution (LR) omnidirectional images (ODIs) to high-resolution (HR), catering to the growing demand for detailed visual content across a $ 180^{\circ}\times360^{\circ}$…

Image and Video Processing · Electrical Eng. & Systems 2026-03-04 Xuhan Sheng , Runyi Li , Bin Chen , Weiqi Li , Xu Jiang , Jian Zhang

In this work, we propose "tangent images," a spherical image representation that facilitates transferable and scalable $360^\circ$ computer vision. Inspired by techniques in cartography and computer graphics, we render a spherical image to…

Computer Vision and Pattern Recognition · Computer Science 2020-05-25 Marc Eder , Mykhailo Shvets , John Lim , Jan-Michael Frahm

Point clouds are a basic data type that is increasingly of interest as 3D content becomes more ubiquitous. Applications using point clouds include virtual, augmented, and mixed reality and autonomous driving. We propose a more efficient…

Computer Vision and Pattern Recognition · Computer Science 2021-06-04 Ryan Killea , Yun Li , Saeed Bastani , Paul McLachlan

As an important technology in 3D mapping, autonomous driving, and robot navigation, LiDAR odometry is still a challenging task. Appropriate data structure and unsupervised deep learning are the keys to achieve an easy adjusted LiDAR…

Computer Vision and Pattern Recognition · Computer Science 2020-11-03 Deyu Yin , Qian Zhang , Jingbin Liu , Xinlian Liang , Yunsheng Wang , Jyri Maanpää , Hao Ma , Juha Hyyppä , Ruizhi Chen

In recent years, the task of learned point cloud compression has gained prominence. An important type of point cloud, the spinning LiDAR point cloud, is generated by spinning LiDAR on vehicles. This process results in numerous circular…

Computer Vision and Pattern Recognition · Computer Science 2024-02-09 Ao Luo , Linxin Song , Keisuke Nonaka , Kyohei Unno , Heming Sun , Masayuki Goto , Jiro Katto

Visual odometry is an essential key for a localization module in SLAM systems. However, previous methods require tuning the system to adapt environment changes. In this paper, we propose a learning-based approach for frame-to-frame…

Computer Vision and Pattern Recognition · Computer Science 2020-01-08 Joosung Lee , Sangwon Hwang , Kyungjae Lee , Woo Jin Kim , Junhyeop Lee , Tae-young Chung , Sangyoun Lee

We introduce Low-Shot Open-Set Domain Generalization (LSOSDG), a novel paradigm unifying low-shot learning with open-set domain generalization (ODG). While prompt-based methods using models like CLIP have advanced DG, they falter in…

Computer Vision and Pattern Recognition · Computer Science 2025-03-21 Mohamad Hassan N C , Divyam Gupta , Mainak Singha , Sai Bhargav Rongali , Ankit Jha , Muhammad Haris Khan , Biplab Banerjee

Self-supervised learning (SSL) has emerged as a powerful technique for learning visual representations. While recent SSL approaches achieve strong results in global image understanding, they are limited in capturing the structured…

Computer Vision and Pattern Recognition · Computer Science 2025-08-28 Oussama Hadjerci , Antoine Letienne , Mohamed Abbas Hedjazi , Adel Hafiane

Nonsmooth composite optimization with orthogonality constraints has a wide range of applications in statistical learning and data science. However, this problem is challenging due to its nonsmooth objective and computationally expensive…

Optimization and Control · Mathematics 2026-05-15 Ganzhao Yuan

The framework of dominant learned video compression methods is usually composed of motion prediction modules as well as motion vector and residual image compression modules, suffering from its complex structure and error propagation…

Image and Video Processing · Electrical Eng. & Systems 2021-04-14 Zhenhong Sun , Zhiyu Tan , Xiuyu Sun , Fangyi Zhang , Dongyang Li , Yichen Qian , Hao Li

Omnidirectional vision is becoming increasingly relevant as more efficient $360^o$ image acquisition is now possible. However, the lack of annotated $360^o$ datasets has hindered the application of deep learning techniques on spherical…

Computer Vision and Pattern Recognition · Computer Science 2019-09-17 Antonis Karakottas , Nikolaos Zioulis , Stamatis Samaras , Dimitrios Ataloglou , Vasileios Gkitsas , Dimitrios Zarpalas , Petros Daras

Despite the widespread adoption of vision sensors in edge applications, such as surveillance, the transmission of video data consumes substantial spectrum resources. Semantic communication (SC) offers a solution by extracting and…

Computer Vision and Pattern Recognition · Computer Science 2026-01-06 Yubo Peng , Luping Xiang , Kun Yang , Kezhi Wang , Merouane Debbah

It has recently been demonstrated that spatial resolution adaptation can be integrated within video compression to improve overall coding performance by spatially down-sampling before encoding and super-resolving at the decoder. Significant…

Image and Video Processing · Electrical Eng. & Systems 2021-01-21 Di Ma , Fan Zhang , David R. Bull