English
Related papers

Related papers: NICE: CVPR 2023 Challenge on Zero-shot Image Capti…

200 papers

Cross-Domain Image Retrieval (CDIR) is a challenging task in computer vision, aiming to match images across different visual domains such as sketches, paintings, and photographs. Existing CDIR methods rely either on supervised learning with…

Computer Vision and Pattern Recognition · Computer Science 2026-04-09 Lucas Iijima , Nikolaos Giakoumoglou , Tania Stathaki

This paper reviews the NTIRE 2025 RAW Image Restoration and Super-Resolution Challenge, highlighting the proposed solutions and results. New methods for RAW Restoration and Super-Resolution could be essential in modern Image Signal…

This paper reviews the NTIRE 2024 low light image enhancement challenge, highlighting the proposed solutions and results. The aim of this challenge is to discover an effective network design or solution capable of generating brighter,…

Computer Vision and Pattern Recognition · Computer Science 2024-04-23 Xiaoning Liu , Zongwei Wu , Ao Li , Florin-Alexandru Vasluianu , Yulun Zhang , Shuhang Gu , Le Zhang , Ce Zhu , Radu Timofte , Zhi Jin , Hongjun Wu , Chenxi Wang , Haitao Ling , Yuanhao Cai , Hao Bian , Yuxin Zheng , Jing Lin , Alan Yuille , Ben Shao , Jin Guo , Tianli Liu , Mohao Wu , Yixu Feng , Shuo Hou , Haotian Lin , Yu Zhu , Peng Wu , Wei Dong , Jinqiu Sun , Yanning Zhang , Qingsen Yan , Wenbin Zou , Weipeng Yang , Yunxiang Li , Qiaomu Wei , Tian Ye , Sixiang Chen , Zhao Zhang , Suiyi Zhao , Bo Wang , Yan Luo , Zhichao Zuo , Mingshen Wang , Junhu Wang , Yanyan Wei , Xiaopeng Sun , Yu Gao , Jiancheng Huang , Hongming Chen , Xiang Chen , Hui Tang , Yuanbin Chen , Yuanbo Zhou , Xinwei Dai , Xintao Qiu , Wei Deng , Qinquan Gao , Tong Tong , Mingjia Li , Jin Hu , Xinyu He , Xiaojie Guo , Sabarinathan , K Uma , A Sasithradevi , B Sathya Bama , S. Mohamed Mansoor Roomi , V. Srivatsav , Jinjuan Wang , Long Sun , Qiuying Chen , Jiahong Shao , Yizhi Zhang , Marcos V. Conde , Daniel Feijoo , Juan C. Benito , Alvaro García , Jaeho Lee , Seongwan Kim , Sharif S M A , Nodirkhuja Khujaev , Roman Tsoy , Ali Murtaza , Uswah Khairuddin , Ahmad 'Athif Mohd Faudzi , Sampada Malagi , Amogh Joshi , Nikhil Akalwadi , Chaitra Desai , Ramesh Ashok Tabib , Uma Mudenagudi , Wenyi Lian , Wenjing Lian , Jagadeesh Kalyanshetti , Vijayalaxmi Ashok Aralikatti , Palani Yashaswini , Nitish Upasi , Dikshit Hegde , Ujwala Patil , Sujata C , Xingzhuo Yan , Wei Hao , Minghan Fu , Pooja choksy , Anjali Sarvaiya , Kishor Upla , Kiran Raja , Hailong Yan , Yunkai Zhang , Baiang Li , Jingyi Zhang , Huan Zheng

Understanding the mechanisms underlying human attention is a fundamental challenge for both vision science and artificial intelligence. While numerous computational models of free-viewing have been proposed, less is known about the…

Computer Vision and Pattern Recognition · Computer Science 2023-05-24 Dario Zanca , Andrea Zugarini , Simon Dietz , Thomas R. Altstidl , Mark A. Turban Ndjeuha , Leo Schwinn , Bjoern Eskofier

Autonomous driving without high-definition (HD) maps demands a higher level of active scene understanding. In this competition, the organizers provided the multi-perspective camera images and standard-definition (SD) maps to explore the…

Computer Vision and Pattern Recognition · Computer Science 2024-06-17 Zhongyu Yang , Mai Liu , Jinluo Xie , Yueming Zhang , Chen Shen , Wei Shao , Jichao Jiao , Tengfei Xing , Runbo Hu , Pengfei Xu

Image captioning, which generates natural language descriptions of the visual information in an image, is a crucial task in vision-language research. Previous models have typically addressed this task by aligning the generative capabilities…

Computer Vision and Pattern Recognition · Computer Science 2024-09-02 Qian Cao , Xu Chen , Ruihua Song , Xiting Wang , Xinting Huang , Yuchen Ren

Image captioning, a fundamental task in vision-language understanding, seeks to generate accurate natural language descriptions for provided images. Current image captioning approaches heavily rely on high-quality image-caption pairs, which…

Computer Vision and Pattern Recognition · Computer Science 2023-11-03 Chuanyang Jin

This paper reviews the AIM 2019 challenge on real world super-resolution. It focuses on the participating methods and final results. The challenge addresses the real world setting, where paired true high and low-resolution images are…

The availability of large-scale image captioning and visual question answering datasets has contributed significantly to recent successes in vision-and-language pre-training. However, these datasets are often collected with overrestrictive…

Computer Vision and Pattern Recognition · Computer Science 2021-03-31 Soravit Changpinyo , Piyush Sharma , Nan Ding , Radu Soricut

Given an image and a reference caption, the image caption editing task aims to correct the misalignment errors and generate a refined caption. However, all existing caption editing works are implicit models, ie, they directly produce the…

Computer Vision and Pattern Recognition · Computer Science 2022-07-21 Zhen Wang , Long Chen , Wenbo Ma , Guangxing Han , Yulei Niu , Jian Shao , Jun Xiao

Humans possess the capacity to reason about the future based on a sparse collection of visual cues acquired over time. In order to emulate this ability, we introduce a novel task called Anticipation Captioning, which generates a caption for…

Computer Vision and Pattern Recognition · Computer Science 2023-04-14 Duc Minh Vo , Quoc-An Luong , Akihiro Sugimoto , Hideki Nakayama

AI in dermatology is evolving at a rapid pace but the major limitation to training trustworthy classifiers is the scarcity of data with ground-truth concept level labels, which are meta-labels semantically meaningful to humans. Foundation…

Computer Vision and Pattern Recognition · Computer Science 2024-09-10 Soham Gadgil , Mahtab Bigverdi

The aim of image captioning is to generate textual description of a given image. Though seemingly an easy task for humans, it is challenging for machines as it requires the ability to comprehend the image (computer vision) and consequently…

Computer Vision and Pattern Recognition · Computer Science 2020-11-12 Anubhav Shrimal , Tanmoy Chakraborty

This paper outlines our approach to the 5th CLVision challenge at CVPR, which addresses the Class-Incremental with Repetition (CIR) scenario. In contrast to traditional class incremental learning, this novel setting introduces unique…

Computer Vision and Pattern Recognition · Computer Science 2025-03-21 Panagiota Moraiti , Efstathios Karypidis

Zero-shot Image Captioning (ZIC) increasingly utilizes synthetic datasets generated by text-to-image (T2I) models to mitigate the need for costly manual annotation. However, these T2I models often produce images that exhibit semantic…

Computer Vision and Pattern Recognition · Computer Science 2025-07-25 Si-Woo Kim , MinJu Jeon , Ye-Chan Kim , Soeun Lee , Taewhan Kim , Dong-Jin Kim

Image captioning strives to generate pertinent captions for specified images, situating itself at the crossroads of Computer Vision (CV) and Natural Language Processing (NLP). This endeavor is of paramount importance with far-reaching…

Computer Vision and Pattern Recognition · Computer Science 2024-04-03 Tianrui Liu , Qi Cai , Changxin Xu , Bo Hong , Jize Xiong , Yuxin Qiao , Tsungwei Yang

Image Captioning, the task of automatic generation of image captions, has attracted attentions from researchers in many fields of computer science, being computer vision, natural language processing and machine learning in recent years.…

Computation and Language · Computer Science 2020-02-04 Quan Hoang Lam , Quang Duy Le , Kiet Van Nguyen , Ngan Luu-Thuy Nguyen

We introduce a novel cross-reference image quality assessment method that effectively fills the gap in the image assessment landscape, complementing the array of established evaluation schemes -- ranging from full-reference metrics like…

Computer Vision and Pattern Recognition · Computer Science 2024-07-24 Zirui Wang , Wenjing Bian , Victor Adrian Prisacariu

In this paper, we propose QACE, a new metric based on Question Answering for Caption Evaluation. QACE generates questions on the evaluated caption and checks its content by asking the questions on either the reference caption or the source…

Computation and Language · Computer Science 2021-08-31 Hwanhee Lee , Thomas Scialom , Seunghyun Yoon , Franck Dernoncourt , Kyomin Jung