English
Related papers

Related papers: Beyond MOS: Subjective Image Quality Score Preproc…

200 papers

Image captioning models require the high-level generalization ability to describe the contents of various images in words. Most existing approaches treat the image-caption pairs equally in their training without considering the differences…

Computer Vision and Pattern Recognition · Computer Science 2022-12-15 Hongkuan Zhang , Saku Sugawara , Akiko Aizawa , Lei Zhou , Ryohei Sasano , Koichi Takeda

With the development of image recovery models,especially those based on adversarial and perceptual losses,the detailed texture portions of images are being recovered more naturally.However,these restored images are similar but not identical…

Image and Video Processing · Electrical Eng. & Systems 2022-11-29 Zhiyu Wang , Jiayan Zhuang , Ningyuan Xu , Sichao Ye , Jiangjian Xiao , Chengbin Peng

Image super-resolution (SR) methods typically model degradation to improve reconstruction accuracy in complex and unknown degradation scenarios. However, extracting degradation information from low-resolution images is challenging, which…

Computer Vision and Pattern Recognition · Computer Science 2025-03-12 Zheng Chen , Yulun Zhang , Jinjin Gu , Xin Yuan , Linghe Kong , Guihai Chen , Xiaokang Yang

Automatically learned quality assessment for images has recently become a hot topic due to its usefulness in a wide variety of applications such as evaluating image capture pipelines, storage techniques and sharing media. Despite the…

Computer Vision and Pattern Recognition · Computer Science 2018-07-04 Hossein Talebi , Peyman Milanfar

Instance features in images exhibit spurious correlations with background features, affecting the training process of deep neural classifiers. This leads to insufficient attention to instance features by the classifier, resulting in…

Computer Vision and Pattern Recognition · Computer Science 2024-12-30 Xuewei Li , Zhenzhen Nie , Mei Yu , Zijian Zhang , Jie Gao , Tianyi Xu , Zhiqiang Liu

Data augmentation is a widely used strategy for training robust machine learning models. It partially alleviates the problem of limited data for tasks like speech emotion recognition (SER), where collecting data is expensive and…

Speech Emotion Recognition (SER) systems rely on speech input and emotional labels annotated by humans. However, various emotion databases collect perceptional evaluations in different ways. For instance, the IEMOCAP dataset uses video…

Audio and Speech Processing · Electrical Eng. & Systems 2025-10-15 Huang-Cheng Chou , Haibin Wu , Hung-yi Lee , Chi-Chun Lee

In most practical situations, the compression or transmission of images and videos creates distortions that will eventually be perceived by a human observer. Vice versa, image and video restoration techniques, such as inpainting or…

Computer Vision and Pattern Recognition · Computer Science 2017-11-29 Rafael Reisenhofer , Sebastian Bosse , Gitta Kutyniok , Thomas Wiegand

Recent advances in text-driven image editing have been significant, yet the task of accurately evaluating these edited images continues to pose a considerable challenge. Different from the assessment of text-driven image generation,…

Computer Vision and Pattern Recognition · Computer Science 2025-01-20 Shangkun Sun , Bowen Qu , Xiaoyu Liang , Songlin Fan , Wei Gao

Over the years, various algorithms were developed, attempting to imitate the Human Visual System (HVS), and evaluate the perceptual image quality. However, for certain image distortions, the functionality of the HVS continues to be an…

Image and Video Processing · Electrical Eng. & Systems 2022-08-09 Shira Faigenbaum-Golovin , Or Shimshi

Low-light images suffer from severe noise and low illumination. Current deep learning models that are trained with real-world images have excellent noise reduction, but a ratio parameter must be chosen manually to complete the enhancement…

Image and Video Processing · Electrical Eng. & Systems 2020-04-23 Qingxu Fu , Xiaoguang Di , Yu Zhang

Score-based diffusion models have significantly advanced generative deep learning for image processing. Measurement conditioned models have also been applied to inverse problems such as CT reconstruction. However, the conventional approach,…

Medical Physics · Physics 2025-02-24 Matthew Tivnan , Dufan Wu , Quanzheng Li

Perceptually-inspired objective functions such as the perceptual evaluation of speech quality (PESQ), signal-to-distortion ratio (SDR), and short-time objective intelligibility (STOI), have recently been used to optimize performance of…

Audio and Speech Processing · Electrical Eng. & Systems 2023-03-27 Khandokar Md. Nayem , Donald S. Williamson

Priors are essential for reconstructing images from noisy and/or incomplete measurements. The choice of the prior determines both the quality and uncertainty of recovered images. We propose turning score-based diffusion models into…

Computer Vision and Pattern Recognition · Computer Science 2023-08-30 Berthy T. Feng , Jamie Smith , Michael Rubinstein , Huiwen Chang , Katherine L. Bouman , William T. Freeman

Mean squared error (MSE) and $\ell_p$ norms have largely dominated the measurement of loss in neural networks due to their simplicity and analytical properties. However, when used to assess visual information loss, these simple norms are…

Image and Video Processing · Electrical Eng. & Systems 2020-07-13 Li-Heng Chen , Christos G. Bampis , Zhi Li , Andrey Norkin , Alan C. Bovik

In cinematography, visual attributes such as color grading, contrast, and brightness are manipulated to reinforce the emotional narrative of a scene. However, conventional Image Signal Processors (ISPs) prioritize scene fidelity,…

Image and Video Processing · Electrical Eng. & Systems 2026-05-06 Dor Barber , Rony Zatzarinni , Hava Matichin , Noam Levy

Implicit Sentiment Analysis (ISA) is a crucial research area in natural language processing. Inspired by the idea of large language model Chain of Thought (CoT), this paper introduces a Sentiment Analysis of Thinking (SAoT) framework. The…

Computation and Language · Computer Science 2024-08-23 Zhihua Duan , Jialin Wang

Bias-transforming methods of fairness-aware machine learning aim to correct a non-neutral status quo with respect to a protected attribute (PA). Current methods, however, lack an explicit formulation of what drives non-neutrality. We…

Machine Learning · Computer Science 2025-02-04 Ludwig Bothmann , Philip A. Boustani , Jose M. Alvarez , Giuseppe Casalicchio , Bernd Bischl , Susanne Dandl

Compressive sensing (CS) works to acquire measurements at sub-Nyquist rate and recover the scene images. Existing CS methods always recover the scene images in pixel level. This causes the smoothness of recovered images and lack of…

Computer Vision and Pattern Recognition · Computer Science 2018-11-29 Jiang Du , Xuemei Xie , Chenye Wang , Guangming Shi

Single-image super-resolution (SR) has achieved remarkable progress with deep learning, yet most approaches rely on distortion-oriented losses or heuristic perceptual priors, which often lead to a trade-off between fidelity and visual…

Image and Video Processing · Electrical Eng. & Systems 2026-02-26 Wei Zhou , Yixiao Li , Hadi Amirpour , Xiaoshuai Hao , Jiang Liu , Peng Wang , Hantao Liu