中文
相关论文

相关论文: Results and findings of the 2021 Image Similarity …

200 篇论文

Document Image Machine Translation (DIMT) seeks to translate text embedded in document images from one language to another by jointly modeling both textual content and page layout, bridging optical character recognition (OCR) and natural…

计算机视觉与模式识别 · 计算机科学 2026-03-11 Yaping Zhang , Yupu Liang , Zhiyang Zhang , Zhiyuan Chen , Lu Xiang , Yang Zhao , Yu Zhou , Chengqing Zong

Existing image-to-image transformation approaches primarily focus on synthesizing visually pleasing data. Generating images with correct identity labels is challenging yet much less explored. It is even more challenging to deal with image…

计算机视觉与模式识别 · 计算机科学 2020-06-16 Wei Xiong , Yutong He , Yixuan Zhang , Wenhan Luo , Lin Ma , Jiebo Luo

This paper reviews the NTIRE 2024 low light image enhancement challenge, highlighting the proposed solutions and results. The aim of this challenge is to discover an effective network design or solution capable of generating brighter,…

计算机视觉与模式识别 · 计算机科学 2024-04-23 Xiaoning Liu , Zongwei Wu , Ao Li , Florin-Alexandru Vasluianu , Yulun Zhang , Shuhang Gu , Le Zhang , Ce Zhu , Radu Timofte , Zhi Jin , Hongjun Wu , Chenxi Wang , Haitao Ling , Yuanhao Cai , Hao Bian , Yuxin Zheng , Jing Lin , Alan Yuille , Ben Shao , Jin Guo , Tianli Liu , Mohao Wu , Yixu Feng , Shuo Hou , Haotian Lin , Yu Zhu , Peng Wu , Wei Dong , Jinqiu Sun , Yanning Zhang , Qingsen Yan , Wenbin Zou , Weipeng Yang , Yunxiang Li , Qiaomu Wei , Tian Ye , Sixiang Chen , Zhao Zhang , Suiyi Zhao , Bo Wang , Yan Luo , Zhichao Zuo , Mingshen Wang , Junhu Wang , Yanyan Wei , Xiaopeng Sun , Yu Gao , Jiancheng Huang , Hongming Chen , Xiang Chen , Hui Tang , Yuanbin Chen , Yuanbo Zhou , Xinwei Dai , Xintao Qiu , Wei Deng , Qinquan Gao , Tong Tong , Mingjia Li , Jin Hu , Xinyu He , Xiaojie Guo , Sabarinathan , K Uma , A Sasithradevi , B Sathya Bama , S. Mohamed Mansoor Roomi , V. Srivatsav , Jinjuan Wang , Long Sun , Qiuying Chen , Jiahong Shao , Yizhi Zhang , Marcos V. Conde , Daniel Feijoo , Juan C. Benito , Alvaro García , Jaeho Lee , Seongwan Kim , Sharif S M A , Nodirkhuja Khujaev , Roman Tsoy , Ali Murtaza , Uswah Khairuddin , Ahmad 'Athif Mohd Faudzi , Sampada Malagi , Amogh Joshi , Nikhil Akalwadi , Chaitra Desai , Ramesh Ashok Tabib , Uma Mudenagudi , Wenyi Lian , Wenjing Lian , Jagadeesh Kalyanshetti , Vijayalaxmi Ashok Aralikatti , Palani Yashaswini , Nitish Upasi , Dikshit Hegde , Ujwala Patil , Sujata C , Xingzhuo Yan , Wei Hao , Minghan Fu , Pooja choksy , Anjali Sarvaiya , Kishor Upla , Kiran Raja , Hailong Yan , Yunkai Zhang , Baiang Li , Jingyi Zhang , Huan Zheng

We introduce a new large-scale dataset that links the assessment of image quality issues to two practical vision tasks: image captioning and visual question answering. First, we identify for 39,181 images taken by people who are blind…

计算机视觉与模式识别 · 计算机科学 2020-03-31 Tai-Yin Chiu , Yinan Zhao , Danna Gurari

The advancement of imaging devices and countless images generated everyday pose an increasingly high demand on image denoising, which still remains a challenging task in terms of both effectiveness and efficiency. To improve denoising…

图像与视频处理 · 电气工程与系统科学 2023-05-10 Zhaoming Kong , Fangxi Deng , Haomin Zhuang , Jun Yu , Lifang He , Xiaowei Yang

We present a novel method for image anomaly detection, where algorithms that use samples drawn from some distribution of "normal" data, aim to detect out-of-distribution (abnormal) samples. Our approach includes a combination of encoder and…

图像与视频处理 · 电气工程与系统科学 2020-03-02 Nina Tuluptceva , Bart Bakker , Irina Fedulova , Anton Konushin

Text Detection and recognition is a one of the important aspect of image processing. This paper analyzes and compares the methods to handle this task. It summarizes the fundamental problems and enumerates factors that need consideration…

计算机视觉与模式识别 · 计算机科学 2018-05-03 Tanvi Goswami , Zankhana Barad , Prof. Nikita P. Desai

The advent of the internet, followed shortly by the social media made it ubiquitous in consuming and sharing information between anyone with access to it. The evolution in the consumption of media driven by this change, led to the emergence…

计算机视觉与模式识别 · 计算机科学 2022-06-02 Cyril Vallez , Andrei Kucharavy , Ljiljana Dolamic

Image captioning models require the high-level generalization ability to describe the contents of various images in words. Most existing approaches treat the image-caption pairs equally in their training without considering the differences…

计算机视觉与模式识别 · 计算机科学 2022-12-15 Hongkuan Zhang , Saku Sugawara , Akiko Aizawa , Lei Zhou , Ryohei Sasano , Koichi Takeda

We introduce the novel problem of identifying the photographer behind a photograph. To explore the feasibility of current computer vision techniques to address this problem, we created a new dataset of over 180,000 images taken by 41…

计算机视觉与模式识别 · 计算机科学 2016-06-02 Christopher Thomas , Adriana Kovashka

This paper provides a review of the NTIRE 2026 challenge on mobile real-world image super-resolution, highlighting the proposed solutions and the resulting outcomes. The challenge aims to recover high-resolution (HR) images from…

Change detection plays an important role in most video-based applications. The first stage is to build appropriate background model, which is now becoming increasingly complex as more sophisticated statistical approaches are introduced to…

计算机视觉与模式识别 · 计算机科学 2014-05-27 Dong Liang , Shun'ichi Kaneko

Dense captioning is a newly emerging computer vision topic for understanding images with dense language descriptions. The goal is to densely detect visual concepts (e.g., objects, object parts, and interactions between them) from images,…

计算机视觉与模式识别 · 计算机科学 2017-08-09 Linjie Yang , Kevin Tang , Jianchao Yang , Li-Jia Li

This paper reviews the NTIRE 2020 challenge on real image denoising with focus on the newly introduced dataset, the proposed methods and their results. The challenge is a new version of the previous NTIRE 2019 challenge on real image…

Image prediction methods often struggle on tasks that require changing the positions of objects, such as video prediction, producing blurry images that average over the many positions that objects might occupy. In this paper, we propose a…

计算机视觉与模式识别 · 计算机科学 2022-04-04 Daniel Geng , Max Hamilton , Andrew Owens

In an automated search system, similarity is a key concept in solving a human task. Indeed, human process is usually a natural categorization that underlies many natural abilities such as image recovery, language comprehension, decision…

计算机视觉与模式识别 · 计算机科学 2018-12-19 Yosr Ghozzi , Nesrine Baklouti , Hani Hagras , Mounir Ben Ayed , Adel M. Alimi

Perceptual judgment of image similarity by humans relies on rich internal representations ranging from low-level features to high-level concepts, scene properties and even cultural associations. However, existing methods and datasets…

计算机视觉与模式识别 · 计算机科学 2018-10-22 Amir Rosenfeld , Markus D. Solbach , John K. Tsotsos

Test sets are an integral part of evaluating models and gauging progress in object recognition, and more broadly in computer vision and AI. Existing test sets for object recognition, however, suffer from shortcomings such as bias towards…

计算机视觉与模式识别 · 计算机科学 2023-01-31 Ali Borji

In this paper, we explore and compare multiple solutions to the problem of data augmentation in image classification. Previous work has demonstrated the effectiveness of data augmentation through simple techniques, such as cropping,…

计算机视觉与模式识别 · 计算机科学 2017-12-14 Luis Perez , Jason Wang