English
Related papers

Related papers: Beyond Bj{\o}ntegaard: Limits of Video Compression…

200 papers

Multimodal Large Language Models (MLLMs) have revolutionized video understanding, yet are still limited by context length when processing long videos. Recent methods compress videos by leveraging visual redundancy uniformly, yielding…

Computer Vision and Pattern Recognition · Computer Science 2025-06-10 Xiao Wang , Qingyi Si , Jianlong Wu , Shiyu Zhu , Li Cao , Liqiang Nie

We propose a perceptual video quality assessment (PVQA) metric for distorted videos by analyzing the power spectral density (PSD) of a group of pictures. This is an estimation approach that relies on the changes in video dynamic calculated…

Computer Vision and Pattern Recognition · Computer Science 2018-12-14 Mohammed A. Aabed , Gukyeong Kwon , Ghassan AlRegib

The recent rise in interest in point clouds as an imaging modality has motivated standardization groups such as MPEG and JPEG Pleno to launch activities aiming at developing compression standards for point clouds. Lossy compression usually…

Image and Video Processing · Electrical Eng. & Systems 2024-11-01 Davi Lazzarotto , Michela Testolina , Touradj Ebrahimi

Learning-based image compression methods have recently emerged as promising alternatives to traditional codecs, offering improved rate-distortion performance and perceptual quality. JPEG AI represents the latest standardized framework in…

Image and Video Processing · Electrical Eng. & Systems 2025-04-11 Mohsen Jenadeleh , Jon Sneyers , Panqi Jia , Shima Mohammadi , Joao Ascenso , Dietmar Saupe

User generated content (UGC) refers to videos that are uploaded by users and shared over the Internet. UGC may have low quality due to noise and previous compression. When re-encoding UGC for streaming or downloading, a traditional video…

Image and Video Processing · Electrical Eng. & Systems 2023-03-14 Xin Xiong , Eduardo Pavez , Antonio Ortega , Balu Adsumilli

This paper describes a quality assessment model for perceptual video compression applications (PVM), which stimulates visual masking and distortion-artefact perception using an adaptive combination of noticeable distortions and blurring…

Image and Video Processing · Electrical Eng. & Systems 2021-06-16 Fan Zhang , David R. Bull

The quality of visual input is very important for both human and machine perception. Consequently many processing techniques exist that deal with different distortions. Usually image processing is applied freely and lacks redundancy…

Image and Video Processing · Electrical Eng. & Systems 2021-02-10 Sascha Xu , Jan Bauer , Benjamin Axmann

Video frame interpolation (VFI) is one of the fundamental research areas in video processing and there has been extensive research on novel and enhanced interpolation algorithms. The same is not true for quality assessment of the…

Image and Video Processing · Electrical Eng. & Systems 2024-11-22 Duolikun Danier , Fan Zhang , David Bull

Standard video frame interpolation methods first estimate optical flow between input frames and then synthesize an intermediate frame guided by motion. Recent approaches merge these two steps into a single convolution process by convolving…

Computer Vision and Pattern Recognition · Computer Science 2017-08-08 Simon Niklaus , Long Mai , Feng Liu

Bit depth adaptation, where the bit depth of a video sequence is reduced before transmission and up-sampled during display, can potentially reduce data rates with limited impact on perceptual quality. In this context, we conducted a…

Image and Video Processing · Electrical Eng. & Systems 2021-09-17 Alex Mackin , Di Ma , Fan Zhang , David Bull

A modular method was suggested before to recover a band limited signal from the sample and hold and linearly interpolated (or, in general, an nth-order-hold) version of the regular samples. In this paper a novel approach for compensating…

Computer Vision and Pattern Recognition · Computer Science 2012-05-15 Mohammad Tofighi , Ali Ayremlou , Farokh Marvasti

Motivated by the work of Uehara et al. [1], an improved method to recover DC coefficients from AC coefficients of DCT-transformed images is investigated in this work, which finds applications in cryptanalysis of selective multimedia…

Multimedia · Computer Science 2015-03-13 Shujun Li , Junaid Jameel Ahmad , Dietmar Saupe , C. -C. Jay Kuo

Multiple beyond-CMOS information processing devices are presently under active research and require methods of benchmarking them. A new approach for calculating the performance metric, energy-delay product, of such devices is proposed. The…

Mesoscale and Nanoscale Physics · Physics 2015-06-17 Angik Sarkar , Dmitri E. Nikonov , Ian A. Young , Behtash Behin-Aein , Supriyo Datta

The AOMedia Video 1 (AV1) standard can achieve considerable compression efficiency thanks to the usage of many advanced tools and improvements, such as advanced inter-prediction modes. However, these come at the cost of high computational…

Image and Video Processing · Electrical Eng. & Systems 2019-08-30 Jieon Kim , Saverio Blasi , Andre Seixas Dias , Marta Mrak , Ebroul Izquierdo

Video retrieval is becoming increasingly important owing to the rapid emergence of videos on the Internet. The dominant paradigm for video retrieval learns video-text representations by pushing the distance between the similarity of…

Computer Vision and Pattern Recognition · Computer Science 2023-03-10 Feng He , Qi Wang , Zhifan Feng , Wenbin Jiang , Yajuan Lv , Yong zhu , Xiao Tan

A concatenated coding scheme over binary memoryless symmetric (BMS) channels using a polarization transformation followed by outer sub-codes is analyzed. Achievable error exponents and upper bounds on the error rate are derived. The first…

Information Theory · Computer Science 2017-10-24 Dina Goldin , David Burshtein

With a plethora of available classification performance measures, choosing the right metric for the right task requires careful thought. To make this decision in an informed manner, one should study and compare general properties of…

Other Computer Science · Computer Science 2020-07-30 Dariusz Brzezinski , Jerzy Stefanowski , Robert Susmaga , Izabela Szczęch

Intra prediction is a crucial component in traditional video coding frameworks, aiming to eliminate spatial redundancy within frames. In recent years, an increasing number of decoder-side adaptive mode derivation methods have been adopted…

Image and Video Processing · Electrical Eng. & Systems 2025-10-01 Jiaqi Zhang , Jiaye Fu , Chuanmin Jia , Siwei Ma , Karam Naser , Thierry Dumas , Saurabh Puri , Milos Radosavljevic

Geometry-based point cloud compression (G-PCC), an international standard designed by MPEG, provides a generic framework for compressing diverse types of point clouds while ensuring interoperability across applications and devices. However,…

Computer Vision and Pattern Recognition · Computer Science 2026-01-16 Wanhao Ma , Wei Zhang , Shuai Wan , Fuzheng Yang

Traditional approaches to interpolate/extrapolate frames in a video sequence require accurate pixel correspondences between images, e.g., using optical flow. Their results stem on the accuracy of optical flow estimation, and could generate…

Computer Vision and Pattern Recognition · Computer Science 2018-03-21 Zhe Hu , Yinglan Ma , Lizhuang Ma
‹ Prev 1 8 9 10 Next ›