English
Related papers

Related papers: AIM 2024 Challenge on Video Saliency Prediction: M…

200 papers

This paper reviews the NTIRE 2024 low light image enhancement challenge, highlighting the proposed solutions and results. The aim of this challenge is to discover an effective network design or solution capable of generating brighter,…

Computer Vision and Pattern Recognition · Computer Science 2024-04-23 Xiaoning Liu , Zongwei Wu , Ao Li , Florin-Alexandru Vasluianu , Yulun Zhang , Shuhang Gu , Le Zhang , Ce Zhu , Radu Timofte , Zhi Jin , Hongjun Wu , Chenxi Wang , Haitao Ling , Yuanhao Cai , Hao Bian , Yuxin Zheng , Jing Lin , Alan Yuille , Ben Shao , Jin Guo , Tianli Liu , Mohao Wu , Yixu Feng , Shuo Hou , Haotian Lin , Yu Zhu , Peng Wu , Wei Dong , Jinqiu Sun , Yanning Zhang , Qingsen Yan , Wenbin Zou , Weipeng Yang , Yunxiang Li , Qiaomu Wei , Tian Ye , Sixiang Chen , Zhao Zhang , Suiyi Zhao , Bo Wang , Yan Luo , Zhichao Zuo , Mingshen Wang , Junhu Wang , Yanyan Wei , Xiaopeng Sun , Yu Gao , Jiancheng Huang , Hongming Chen , Xiang Chen , Hui Tang , Yuanbin Chen , Yuanbo Zhou , Xinwei Dai , Xintao Qiu , Wei Deng , Qinquan Gao , Tong Tong , Mingjia Li , Jin Hu , Xinyu He , Xiaojie Guo , Sabarinathan , K Uma , A Sasithradevi , B Sathya Bama , S. Mohamed Mansoor Roomi , V. Srivatsav , Jinjuan Wang , Long Sun , Qiuying Chen , Jiahong Shao , Yizhi Zhang , Marcos V. Conde , Daniel Feijoo , Juan C. Benito , Alvaro García , Jaeho Lee , Seongwan Kim , Sharif S M A , Nodirkhuja Khujaev , Roman Tsoy , Ali Murtaza , Uswah Khairuddin , Ahmad 'Athif Mohd Faudzi , Sampada Malagi , Amogh Joshi , Nikhil Akalwadi , Chaitra Desai , Ramesh Ashok Tabib , Uma Mudenagudi , Wenyi Lian , Wenjing Lian , Jagadeesh Kalyanshetti , Vijayalaxmi Ashok Aralikatti , Palani Yashaswini , Nitish Upasi , Dikshit Hegde , Ujwala Patil , Sujata C , Xingzhuo Yan , Wei Hao , Minghan Fu , Pooja choksy , Anjali Sarvaiya , Kishor Upla , Kiran Raja , Hailong Yan , Yunkai Zhang , Baiang Li , Jingyi Zhang , Huan Zheng

Computational models of visual attention in artificial intelligence and robotics have been inspired by the concept of a saliency map. These models account for the mutual information between the (current) visual information and its estimated…

Robotics · Computer Science 2022-03-25 Ajith Anil Meera , Filip Novicky , Thomas Parr , Karl Friston , Pablo Lanillos , Noor Sajid

This paper presents a summary of the VQualA 2025 Challenge on Visual Quality Comparison for Large Multimodal Models (LMMs), hosted as part of the ICCV 2025 Workshop on Visual Quality Assessment. The challenge aims to evaluate and enhance…

AutoFocus-IL is a simple yet effective method to improve data efficiency and generalization in visual imitation learning by guiding policies to attend to task-relevant features rather than distractors and spurious correlations. Although…

Robotics · Computer Science 2025-11-26 Litian Gong , Fatemeh Bahrani , Yutai Zhou , Amin Banayeeanzade , Jiachen Li , Erdem Bıyık

In this paper we propose a Kalman filter aided saliency detection model which is based on the conjecture that salient regions are considerably different from our "visual expectation" or they are "visually surprising" in nature. In this…

Computer Vision and Pattern Recognition · Computer Science 2016-04-19 Sourya Roy , Pabitra Mitra

Human visual attention is a complex phenomenon. A computational modeling of this phenomenon must take into account where people look in order to evaluate which are the salient locations (spatial distribution of the fixations), when they…

Computer Vision and Pattern Recognition · Computer Science 2020-05-08 Dario Zanca , Stefano Melacci , Marco Gori

Automatic video summarization is still an unsolved problem due to several challenges. We take steps towards making automatic video summarization more realistic by addressing them. Firstly, the currently available datasets either have very…

Computer Vision and Pattern Recognition · Computer Science 2020-08-26 Vishal Kaushal , Suraj Kothawade , Rishabh Iyer , Ganesh Ramakrishnan

The AI City Challenge was created to accelerate intelligent video analysis that helps make cities smarter and safer. Transportation is one of the largest segments that can benefit from actionable insights derived from data captured by…

Computer Vision and Pattern Recognition · Computer Science 2020-05-01 Milind Naphade , Shuo Wang , David Anastasiu , Zheng Tang , Ming-Ching Chang , Xiaodong Yang , Liang Zheng , Anuj Sharma , Rama Chellappa , Pranamesh Chakraborty

This paper presents a review of the LoViF 2026 Challenge on Weather Removal in Videos. The challenge encourages the development of methods for restoring clean videos from inputs degraded by adverse weather conditions such as rain and snow,…

In this paper, we present a solution to Large-Scale Video Classification Challenge (LSVC2017) [1] that ranked the 1st place. We focused on a variety of modalities that cover visual, motion and audio. Also, we visualized the aggregation…

Computer Vision and Pattern Recognition · Computer Science 2017-10-31 Chen Chen , Xiaowei Zhao , Yang Liu

Saliency maps have become a widely used method to make deep learning models more interpretable by providing post-hoc explanations of classifiers through identification of the most pertinent areas of the input medical image. They are…

With the rise of short videos, the demand for selecting appropriate background music (BGM) for a video has increased significantly, video-music retrieval (VMR) task gradually draws much attention by research community. As other cross-modal…

Multimedia · Computer Science 2023-02-21 Xuxin Cheng , Zhihong Zhu , Hongxiang Li , Yaowei Li , Yuexian Zou

Activity detection in surveillance videos is a challenging task caused by small objects, complex activity categories, its untrimmed nature, etc. Existing methods are generally limited in performance due to inaccurate proposals, poor…

Computer Vision and Pattern Recognition · Computer Science 2022-05-10 Yunhao Du , Zhihang Tong , Junfeng Wan , Binyu Zhang , Yanyun Zhao

Visual Saliency refers to the innate human mechanism of focusing on and extracting important features from the observed environment. Recently, there has been a notable surge of interest in the field of automotive research regarding the…

Computer Vision and Pattern Recognition · Computer Science 2023-08-09 Francesco Rundo , Michael Sebastian Rundo , Concetto Spampinato

Textured meshes significantly enhance the realism and detail of objects by mapping intricate texture details onto the geometric structure of 3D models. This advancement is valuable across various applications, including entertainment,…

Graphics · Computer Science 2024-12-12 Kaiwei Zhang , Dandan Zhu , Xiongkuo Min , Guangtao Zhai

This article reports on an investigation of the use of convolutional neural networks to predict the visual attention of chess players. The visual attention model described in this article has been created to generate saliency maps that…

Machine Learning · Statistics 2019-04-21 Justin Le Louedec , Thomas Guntz , James Crowley , Dominique Vaufreydaz

Saliency detection has drawn a lot of attention of researchers in various fields over the past several years. Saliency is the perceptual quality that makes an object, person to draw the attention of humans at the very sight. Salient object…

Computer Vision and Pattern Recognition · Computer Science 2017-07-06 Shubham Pachori

Saliency integration has attracted much attention on unifying saliency maps from multiple saliency models. Previous offline integration methods usually face two challenges: 1. if most of the candidate saliency models misjudge the saliency…

Computer Vision and Pattern Recognition · Computer Science 2018-07-30 Yingyue Xu , Xiaopeng Hong , Fatih Porikli , Xin Liu , Jie Chen , Guoying Zhao

Visual attention modeling, important for interpreting and prioritizing visual stimuli, plays a significant role in applications such as marketing, multimedia, and robotics. Traditional saliency prediction models, especially those based on…

Computer Vision and Pattern Recognition · Computer Science 2024-09-10 Alireza Hosseini , Amirhossein Kazerouni , Saeed Akhavan , Michael Brudno , Babak Taati

Understanding how biological visual systems process information is challenging due to the complex nonlinear relationship between neuronal responses and high-dimensional visual input. Artificial neural networks have already improved our…

‹ Prev 1 8 9 10 Next ›