中文
相关论文

相关论文: Users prefer Guetzli JPEG over same-sized libjpeg

200 篇论文

Interest is increasing among political scientists in leveraging the extensive information available in images. However, the challenge of interpreting these images lies in the need for specialized knowledge in computer vision and access to…

计算机视觉与模式识别 · 计算机科学 2024-03-28 Yu Wang

Qualitative analysis is typically limited to small datasets because it is time-intensive. Moreover, a second human rater is required to ensure reliable findings. Artificial intelligence tools may replace human raters if we demonstrate high…

物理教育 · 物理学 2025-09-03 Nikhil Sanjay Borse , Ravishankar Chatta Subramaniam , N. Sanjay Rebello

Empirical evidence has demonstrated that learning-based image compression can outperform classical compression frameworks. This has led to the ongoing standardization of learned-based image codecs, namely Joint Photographic Experts Group…

图像与视频处理 · 电气工程与系统科学 2025-03-21 Panqi Jia , Fabian Brand , Dequan Yu , Alexander Karabutov , Elena Alshina , Andre Kaup

This paper investigates the causality in the decision making of movie recommendations through the users' affective profiles. We advocate a method of assigning emotional tags to a movie by the auto-detection of the affective features in the…

信息检索 · 计算机科学 2021-02-12 John Kalung Leung , Igor Griva , William G. Kennedy

Text and audio simplification to increase information comprehension are important in healthcare. With the introduction of ChatGPT, an evaluation of its simplification performance is needed. We provide a systematic comparison of human and…

计算与语言 · 计算机科学 2024-05-06 Gondy Leroy , David Kauchak , Philip Harber , Ankit Pal , Akash Shukla

Automated sentiment analysis using Large Language Model (LLM)-based models like ChatGPT, Gemini or LLaMA2 is becoming widespread, both in academic research and in industrial applications. However, assessment and validation of their…

计算与语言 · 计算机科学 2024-02-06 Alessio Buscemi , Daniele Proverbio

Leveraging Large Multimodal Models (LMMs) to simulate human behaviors when processing multimodal information, especially in the context of social media, has garnered immense interest due to its broad potential and far-reaching implications.…

计算机视觉与模式识别 · 计算机科学 2024-04-16 Hanjia Lyu , Weihong Qi , Zhongyu Wei , Jiebo Luo

In the last decade, the use of simple rating and comparison surveys has proliferated on social and digital media platforms to fuel recommendations. These simple surveys and their extrapolation with machine learning algorithms shed light on…

社会与信息网络 · 计算机科学 2019-01-29 Nandana Sengupta , Nati Srebro , James Evans

We consider the task of upscaling a low resolution thumbnail image of a person, to a higher resolution image, which preserves the person's identity and other attributes. Since the thumbnail image is of low resolution, many higher resolution…

计算机视觉与模式识别 · 计算机科学 2021-06-01 Noam Gat , Sagie Benaim , Lior Wolf

Video dimensions are continuously increasing to provide more realistic and immersive experiences to global streaming and social media viewers. However, increments in video parameters such as spatial resolution and frame rate are inevitably…

图像与视频处理 · 电气工程与系统科学 2022-01-19 Dae Yeol Lee , Somdyuti Paul , Christos G. Bampis , Hyunsuk Ko , Jongho Kim , Se Yoon Jeong , Blake Homan , Alan C. Bovik

Advances in automated scoring are closely aligned with advances in machine-learning and natural-language-processing techniques. With recent progress in large language models (LLMs), the use of ChatGPT, Gemini, Claude, and other…

计算与语言 · 计算机科学 2025-09-30 Haowei Hua , Hong Jiao , Dan Song

Selfies have become increasingly fashionable in the social media era. People are willing to share their selfies in various social media platforms such as Facebook, Instagram and Flicker. The popularity of selfie have caught researchers'…

社会与信息网络 · 计算机科学 2017-02-28 Tianlang Chen , Yuxiao Chen , Jiebo Luo

We introduce VisualQuest, a novel dataset designed to rigorously evaluate multimodal large language models (MLLMs) on abstract visual reasoning tasks that require the integration of symbolic, cultural, and linguistic knowledge. Unlike…

计算机视觉与模式识别 · 计算机科学 2026-01-05 Kelaiti Xiao , Liang Yang , Dongyu Zhang , Paerhati Tulajiang , Hongfei Lin

The research on neural network (NN) based image compression has shown superior performance compared to classical compression frameworks. Unlike the hand-engineered transforms in the classical frameworks, NN-based models learn the non-linear…

计算机视觉与模式识别 · 计算机科学 2024-02-28 Panqi Jia , A. Burakhan Koyuncu , Jue Mao , Ze Cui , Yi Ma , Tiansheng Guo , Timofey Solovyev , Alexander Karabutov , Yin Zhao , Jing Wang , Elena Alshina , Andre Kaup

Compression plays a significant role in a data storage and a transmission. If we speak about a generall data compression, it has to be a lossless one. It means, we are able to recover the original data 1:1 from the compressed file.…

图形学 · 计算机科学 2014-10-10 Martin Prantl

This literature has proposed three fast and easy computable image features to improve computer vision by offering more human-like vision power. These features are not based on image pixels absolute or relative intensity; neither based on…

计算机视觉与模式识别 · 计算机科学 2020-04-16 Soumi Ray , Vinod Kumar

Progress in lighting estimation is tracked by computing existing image quality assessment (IQA) metrics on images from standard datasets. While this may appear to be a reasonable approach, we demonstrate that doing so does not correlate to…

计算机视觉与模式识别 · 计算机科学 2024-03-22 Justine Giroux , Mohammad Reza Karimi Dastjerdi , Yannick Hold-Geoffroy , Javier Vazquez-Corral , Jean-François Lalonde

Recent advances in image editing have enabled models to handle complex instructions with impressive realism. However, existing evaluation frameworks lag behind: current benchmarks suffer from narrow task coverage, while standard metrics…

计算机视觉与模式识别 · 计算机科学 2026-03-31 Zhangqi Jiang , Zheng Sun , Xianfang Zeng , Yufeng Yang , Xuanyang Zhang , Yongliang Wu , Wei Cheng , Gang Yu , Xu Yang , Bihan Wen

Most neural networks for computer vision are designed to infer using RGB images. However, these RGB images are commonly encoded in JPEG before saving to disk; decoding them imposes an unavoidable overhead for RGB networks. Instead, our work…

计算机视觉与模式识别 · 计算机科学 2023-06-16 Jeongsoo Park , Justin Johnson

In an automated search system, similarity is a key concept in solving a human task. Indeed, human process is usually a natural categorization that underlies many natural abilities such as image recovery, language comprehension, decision…

计算机视觉与模式识别 · 计算机科学 2018-12-19 Yosr Ghozzi , Nesrine Baklouti , Hani Hagras , Mounir Ben Ayed , Adel M. Alimi