English
Related papers

Related papers: ViSAGE @ NTIRE 2026 Challenge on Video Saliency Pr…

200 papers

Video Super-Resolution (VSR) aims to recover sequences of high-resolution (HR) frames from low-resolution (LR) frames. Previous methods mainly utilize temporally adjacent frames to assist the reconstruction of target frames. However, in the…

Computer Vision and Pattern Recognition · Computer Science 2023-04-12 Yongjie Chen , Tieru Wu

In order to deal with the task of video panoptic segmentation in the wild, we propose a robust integrated video panoptic segmentation solution. In our solution, we regard the video panoptic segmentation task as a segmentation target…

Computer Vision and Pattern Recognition · Computer Science 2023-06-13 Jinming Su , Wangwang Yang , Junfeng Luo , Xiaolin Wei

Computational modeling of visual saliency has become an important research problem in recent years, with applications in video quality estimation, video compression, object tracking, retargeting, summarization, and so on. While most visual…

Multimedia · Computer Science 2016-04-26 Sayed Hossein Khatoonabadi , Ivan V. Bajic , Yufeng Shan

Survival analysis using pathology images poses a considerable challenge, as it requires the localization of relevant information from the multitude of tiles within whole slide images (WSIs). Current methods typically resort to a two-stage…

Image and Video Processing · Electrical Eng. & Systems 2024-09-09 Zhongwei Qiu , Hanqing Chao , Wenbin Liu , Yixuan Shen , Le Lu , Ke Yan , Dakai Jin , Yun Bian , Hui Jiang

With the rapid development of pre-training technologies, adapting large-scale Vision-Language Models (VLMs) for video understanding \emph{\ie} image-to-video transfer learning has become a dominant paradigm. To achieve superior performance,…

Computer Vision and Pattern Recognition · Computer Science 2026-05-28 Rui Lin , Chuanming Wang , Huadong Ma

Video summarization aims to produce a compact representation of a long video by selecting a subset of temporally important segments that best reflect human preferences. This task is inherently difficult due to strong annotation subjectivity…

Computer Vision and Pattern Recognition · Computer Science 2026-05-12 Omer Tariq , Syed Muhammad Raza , Jeongbae Son

Existing state-of-the-art saliency detection methods heavily rely on CNN-based architectures. Alternatively, we rethink this task from a convolution-free sequence-to-sequence perspective and predict saliency by modeling long-range…

Computer Vision and Pattern Recognition · Computer Science 2021-08-24 Nian Liu , Ni Zhang , Kaiyuan Wan , Ling Shao , Junwei Han

We propose a novel mixture-of-experts class to optimize computer vision models in accordance with data transfer limitations at test time. Our approach postulates that the minimum acceptable amount of data allowing for highly-accurate…

Machine Learning · Computer Science 2020-09-02 Alhabib Abbas , Yiannis Andreopoulos

This report describes our solution to the VALUE Challenge 2021 in the captioning task. Our solution, named CLIP4Caption++, is built on X-Linear/X-Transformer, which is an advanced model with encoder-decoder architecture. We make the…

Computer Vision and Pattern Recognition · Computer Science 2021-10-15 Mingkang Tang , Zhanyu Wang , Zhaoyang Zeng , Fengyun Rao , Dian Li

This paper introduces a new framework to predict visual attention of omnidirectional images. The key setup of our architecture is the simultaneous prediction of the saliency map and a corresponding scanpath for a given stimulus. The…

Computer Vision and Pattern Recognition · Computer Science 2022-01-04 Mohamed Amine Kerkouri , Marouane Tliba , Aladine Chetouani , Mohamed Sayeh

Video saliency prediction has recently attracted attention of the research community, as it is an upstream task for several practical applications. However, current solutions are particularly computationally demanding, especially due to the…

Computer Vision and Pattern Recognition · Computer Science 2023-01-12 Feiyan Hu , Simone Palazzo , Federica Proietto Salanitri , Giovanni Bellitto , Morteza Moradi , Concetto Spampinato , Kevin McGuinness

Saliency computation models aim to imitate the attention mechanism in the human visual system. The application of deep neural networks for saliency prediction has led to a drastic improvement over the last few years. However, deep models…

Computer Vision and Pattern Recognition · Computer Science 2024-03-26 Saman Zabihi , Hamed Rezazadegan Tavakoli , Ali Borji

This paper presents a review for the NTIRE 2025 Challenge on Short-form UGC Video Quality Assessment and Enhancement. The challenge comprises two tracks: (i) Efficient Video Quality Assessment (KVQ), and (ii) Diffusion-based Image…

Image and Video Processing · Electrical Eng. & Systems 2025-04-18 Xin Li , Kun Yuan , Bingchen Li , Fengbin Guan , Yizhen Shao , Zihao Yu , Xijun Wang , Yiting Lu , Wei Luo , Suhang Yao , Ming Sun , Chao Zhou , Zhibo Chen , Radu Timofte , Yabin Zhang , Ao-Xiang Zhang , Tianwu Zhi , Jianzhao Liu , Yang Li , Jingwen Xu , Yiting Liao , Yushen Zuo , Mingyang Wu , Renjie Li , Shengyun Zhong , Zhengzhong Tu , Yufan Liu , Xiangguang Chen , Zuowei Cao , Minhao Tang , Shan Liu , Kexin Zhang , Jingfen Xie , Yan Wang , Kai Chen , Shijie Zhao , Yunchen Zhang , Xiangkai Xu , Hong Gao , Ji Shi , Yiming Bao , Xiugang Dong , Xiangsheng Zhou , Yaofeng Tu , Ying Liang , Yiwen Wang , Xinning Chai , Yuxuan Zhang , Zhengxue Cheng , Yingsheng Qin , Yucai Yang , Rong Xie , Li Song , Wei Sun , Kang Fu , Linhan Cao , Dandan Zhu , Kaiwei Zhang , Yucheng Zhu , Zicheng Zhang , Menghan Hu , Xiongkuo Min , Guangtao Zhai , Zhi Jin , Jiawei Wu , Wei Wang , Wenjian Zhang , Yuhai Lan , Gaoxiong Yi , Hengyuan Na , Wang Luo , Di Wu , MingYin Bai , Jiawang Du , Zilong Lu , Zhenyu Jiang , Hui Zeng , Ziguan Cui , Zongliang Gan , Guijin Tang , Xinglin Xie , Kehuan Song , Xiaoqiang Lu , Licheng Jiao , Fang Liu , Xu Liu , Puhua Chen , Ha Thu Nguyen , Katrien De Moor , Seyed Ali Amirshahi , Mohamed-Chaker Larabi , Qi Tang , Linfeng He , Zhiyong Gao , Zixuan Gao , Guohua Zhang , Zhiye Huang , Yi Deng , Qingmiao Jiang , Lu Chen , Yi Yang , Xi Liao , Nourine Mohammed Nadir , Yuxuan Jiang , Qiang Zhu , Siyue Teng , Fan Zhang , Shuyuan Zhu , Bing Zeng , David Bull , Meiqin Liu , Chao Yao , Yao Zhao

Despite significant progress, image saliency detection still remains a challenging task in complex scenes and environments. Integrating multiple different but complementary cues, like RGB and Thermal (RGB-T), may be an effective way for…

Computer Vision and Pattern Recognition · Computer Science 2017-01-12 Chenglong Li , Guizhao Wang , Yunpeng Ma , Aihua Zheng , Bin Luo , Jin Tang

Gradient-based saliency methods such as Vanilla Gradient (VG) and Integrated Gradients (IG) are widely used to explain image classifiers, yet the resulting maps are often noisy and unstable, limiting their usefulness in high-stakes…

Computer Vision and Pattern Recognition · Computer Science 2026-03-10 Dipkamal Bhusal , Md Tanvirul Alam , Nidhi Rastogi

Video Variational Autoencoder (VAE) enables latent video generative modeling by mapping the visual world into compact spatiotemporal latent spaces, improving training efficiency and stability. While existing video VAEs achieve commendable…

Computer Vision and Pattern Recognition · Computer Science 2026-05-05 Yian Zhao , Feng Wang , Qiushan Guo , Chang Liu , Xiangyang Ji , Jian Zhang , Jie Chen

This paper introduces ViNet-S, a 36MB model based on the ViNet architecture with a U-Net design, featuring a lightweight decoder that significantly reduces model size and parameters without compromising performance. Additionally, ViNet-A…

Computer Vision and Pattern Recognition · Computer Science 2025-02-04 Rohit Girmaji , Siddharth Jain , Bhav Beri , Sarthak Bansal , Vineet Gandhi

Video anomaly detection (VAD) is currently a challenging task due to the complexity of anomaly as well as the lack of labor-intensive temporal annotations. In this paper, we propose an end-to-end Global Information Guided (GIG) anomaly…

Computer Vision and Pattern Recognition · Computer Science 2021-04-15 Hui Lv , Chunyan Xu , Zhen Cui

This paper presents an approach for top-down saliency detection guided by visual classification tasks. We first learn how to compute visual saliency when a specific visual task has to be accomplished, as opposed to most state-of-the-art…

Computer Vision and Pattern Recognition · Computer Science 2018-03-28 Francesca Murabito , Concetto Spampinato , Simone Palazzo , Konstantin Pogorelov , Michael Riegler

Recent advances in image-based saliency prediction are approaching gold standard performance levels on existing benchmarks. Despite this success, we show that predicting fixations across multiple saliency datasets remains challenging due to…

Computer Vision and Pattern Recognition · Computer Science 2025-10-01 Matthias Kümmerer , Harneet Singh Khanuja , Matthias Bethge
‹ Prev 1 4 5 6 7 8 10 Next ›