中文
相关论文

相关论文: LoViF 2026 Challenge on Human-oriented Semantic Im…

200 篇论文

Large Vision-Language Models (LVLMs), despite their recent success, are hardly comprehensively tested for their cognitive abilities. Inspired by the prevalent use of the Cookie Theft task in human cognitive tests, we propose a novel…

人工智能 · 计算机科学 2025-02-14 Xiujie Song , Mengyue Wu , Kenny Q. Zhu , Chunhao Zhang , Yanyi Chen

Evaluating text-to-vision content hinges on two crucial aspects: visual quality and alignment. While significant progress has been made in developing objective models to assess these dimensions, the performance of such models heavily relies…

计算机视觉与模式识别 · 计算机科学 2025-06-17 Zicheng Zhang , Tengchuan Kou , Shushi Wang , Chunyi Li , Wei Sun , Wei Wang , Xiaoyu Li , Zongyu Wang , Xuezhi Cao , Xiongkuo Min , Xiaohong Liu , Guangtao Zhai

Image composition involves extracting a foreground object from one image and pasting it into another image through Image harmonization algorithms (IHAs), which aim to adjust the appearance of the foreground object to better match the…

计算机视觉与模式识别 · 计算机科学 2025-01-03 Zitong Xu , Huiyu Duan , Guangji Ma , Liu Yang , Jiarui Wang , Qingbo Wu , Xiongkuo Min , Guangtao Zhai , Patrick Le Callet

This paper presents the NTIRE 2026 image super-resolution ($\times$4) challenge, one of the associated competitions of the NTIRE 2026 Workshop at CVPR 2026. The challenge aims to reconstruct high-resolution (HR) images from low-resolution…

计算机视觉与模式识别 · 计算机科学 2026-04-17 Zheng Chen , Kai Liu , Jingkai Wang , Xianglong Yan , Jianze Li , Ziqing Zhang , Jue Gong , Jiatong Li , Lei Sun , Xiaoyang Liu , Radu Timofte , Yulun Zhang , Jihye Park , Yoonjin Im , Hyungju Chun , Hyunhee Park , MinKyu Park , Zheng Xie , Xiangyu Kong , Weijun Yuan , Zhan Li , Qiurong Song , Luen Zhu , Fengkai Zhang , Xinzhe Zhu , Junyang Chen , Congyu Wang , Yixin Yang , Zhaorun Zhou , Jiangxin Dong , Jinshan Pan , Shengwei Wang , Jiajie Ou , Baiang Li , Sizhuo Ma , Qiang Gao , Jusheng Zhang , Jian Wang , Keze Wang , Yijiao Liu , Yingsi Chen , Hui Li , Yu Wang , Congchao Zhu , Saeed Ahmad , Ik Hyun Lee , Jun Young Park , Ji Hwan Yoon , Kainan Yan , Zian Wang , Weibo Wang , Shihao Zou , Chao Dong , Wei Zhou , Linfeng Li , Jaeseong Lee , Jaeho Chae , Jinwoo Kim , Seonjoo Kim , Yucong Hong , Zhenming Yan , Junye Chen , Ruize Han , Song Wang , Yuxuan Jiang , Chengxi Zeng , Tianhao Peng , Fan Zhang , David Bull , Tongyao Mu , Qiong Cao , Yifan Wang , Youwei Pan , Leilei Cao , Xiaoping Peng , Wei Deng , Yifei Chen , Wenbo Xiong , Xian Hu , Yuxin Zhang , Xiaoyun Cheng , Yang Ji , Zonghao Chen , Zhihao Xue , Junqin Hu , Nihal Kumar , Snehal Singh Tomar , Klaus Mueller , Surya Vashisth , Prateek Shaily , Jayant Kumar , Hardik Sharma , Ashish Negi , Sachin Chaudhary , Akshay Dudhane , Praful Hambarde , Amit Shukla , Shijun Shi , Jiangning Zhang , Yong Liu , Kai Hu , Jing Xu , Xianfang Zeng , Amitesh M , Hariharan S , Chia-Ming Lee , Yu-Fan Lin , Chih-Chung Hsu , Nishalini K , Sreenath K A , Bilel Benjdira , Anas M. Ali , Wadii Boulila , Shuling Zheng , Zhiheng Fu , Feng Zhang , Zhanglu Chen , Boyang Yao , Nikhil Pathak , Aagam Jain , Milan Kumar , Kishor Upla , Vivek Chavda , Sarang N S , Raghavendra Ramachandra , Zhipeng Zhang , Qi Wang , Shiyu Wang , Jiachen Tu , Guoyi Xu , Yaoxin Jiang , Jiajia Liu , Yaokun Shi , Yuqi Li , Chuanguang Yang , Weilun Feng , Zhuzhi Hong , Hao Wu , Junming Liu , Yingli Tian , Amish Bhushan Kulkarni , Tejas R R Shet , Saakshi M Vernekar , Nikhil Akalwadi , Kaushik Mallibhat , Ramesh Ashok Tabib , Uma Mudenagudi , Yuwen Pan , Tianrun Chen , Deyi Ji , Qi Zhu , Lanyun Zhu , Heyan Zhangyi

This report presents the results and findings of the first edition of the Short-Films 20K (SF20K) Competition, held in conjunction with the SLoMO Workshop at ICCV 2025. The competition is designed to advance story-level video understanding…

计算机视觉与模式识别 · 计算机科学 2026-05-05 Ridouane Ghermi , Xi Wang , Vicky Kalogeiton , Ivan Laptev

Human parsing aims to partition humans in image or video into multiple pixel-level semantic parts. In the last decade, it has gained significantly increased interest in the computer vision community and has been utilized in a broad range of…

计算机视觉与模式识别 · 计算机科学 2024-03-15 Lu Yang , Wenhe Jia , Shan Li , Qing Song

Image Quality Assessment (IQA) is a critical task in a wide range of applications but remains challenging due to the subjective nature of human perception and the complexity of real-world image distortions. This study proposes MetaQAP, a…

计算机视觉与模式识别 · 计算机科学 2025-10-17 Nisar Ahmed , Gulshan Saleem , Nazik Alturki , Nada Alasbali

This paper provides a review of the NTIRE 2026 challenge on real-world face restoration, highlighting the proposed solutions and the resulting outcomes. The challenge focuses on generating natural and realistic outputs while maintaining…

Aesthetic assessment of images can be categorized into two main forms: numerical assessment and language assessment. Aesthetics caption of photographs is the only task of aesthetic language assessment that has been addressed. In this paper,…

计算机视觉与模式识别 · 计算机科学 2022-08-12 Xin Jin , Wu Zhou , Xinghui Zhou , Shuai Cui , Le Zhang , Jianwen Lv , Shu Zhao

Image description datasets play a crucial role in the advancement of various applications such as image understanding, text-to-image generation, and text-image retrieval. Currently, image description datasets primarily originate from two…

计算机视觉与模式识别 · 计算机科学 2024-06-12 Renjie Pi , Jianshu Zhang , Jipeng Zhang , Rui Pan , Zhekai Chen , Tong Zhang

While text-to-image (T2I) generation models have achieved remarkable progress in recent years, existing evaluation methodologies for vision-language alignment still struggle with the fine-grained semantic matching. Current approaches based…

计算机视觉与模式识别 · 计算机科学 2025-04-11 Zijian Zhang , Xuhui Zheng , Xuecheng Wu , Chong Peng , Xuezhi Cao

Image quality assessment (IQA) algorithm aims to quantify the human perception of image quality. Unfortunately, there is a performance drop when assessing the distortion images generated by generative adversarial network (GAN) with…

计算机视觉与模式识别 · 计算机科学 2022-04-25 Shanshan Lao , Yuan Gong , Shuwei Shi , Sidi Yang , Tianhe Wu , Jiahao Wang , Weihao Xia , Yujiu Yang

Semantic noise in image classification datasets, where visually similar categories are frequently mislabeled, poses a significant challenge to conventional supervised learning approaches. In this paper, we explore the potential of using…

计算机视觉与模式识别 · 计算机科学 2025-09-05 Yingxuan Li , Jiafeng Mao , Yusuke Matsui

Earth vision research typically focuses on extracting geospatial object locations and categories but neglects the exploration of relations between objects and comprehensive reasoning. Based on city planning needs, we develop a multi-modal…

计算机视觉与模式识别 · 计算机科学 2023-12-20 Junjue Wang , Zhuo Zheng , Zihang Chen , Ailong Ma , Yanfei Zhong

We introduce WearVQA, the first benchmark specifically designed to evaluate the Visual Question Answering (VQA) capabilities of multi-model AI assistant on wearable devices like smart glasses. Unlike prior benchmarks that focus on…

Multi-level deep-features have been driving state-of-the-art methods for aesthetics and image quality assessment (IQA). However, most IQA benchmarks are comprised of artificially distorted images, for which features derived from ImageNet…

图像与视频处理 · 电气工程与系统科学 2020-01-23 Hanhe Lin , Vlad Hosu , Dietmar Saupe

Deep learning has driven remarkable accuracy increases in many computer vision problems. One ongoing challenge is how to achieve the greatest accuracy in cases where training data is limited. A second ongoing challenge is that trained…

计算机视觉与模式识别 · 计算机科学 2021-10-22 Aidan Boyd , Kevin Bowyer , Adam Czajka

Low-level image processing has long been evaluated mainly from the perspective of visual fidelity. However, with the rise of deep learning and generative models, processed images may preserve perceptual quality while altering semantic…

计算机视觉与模式识别 · 计算机科学 2026-04-29 Runjie Wang , Weiling Chen , Tiesong Zhao , Chang Wen Chen

Image quality assessment (IQA) is inherently complex, as it reflects both the quantification and interpretation of perceptual quality rooted in the human visual system. Conventional approaches typically rely on fixed models to output scalar…

计算机视觉与模式识别 · 计算机科学 2025-10-02 Hanwei Zhu , Yu Tian , Keyan Ding , Baoliang Chen , Bolin Chen , Shiqi Wang , Weisi Lin

The Semantic Publishing Challenge series aims at investigating novel approaches for improving scholarly publishing using Linked Data technology. In 2014 we had bootstrapped this effort with a focus on extracting information from…

数字图书馆 · 计算机科学 2015-08-26 Angelo Di Iorio , Christoph Lange , Anastasia Dimou , Sahar Vahdati