中文
相关论文

相关论文: ICDAR 2019 Competition on Image Retrieval for Hist…

200 篇论文

When extracting information from handwritten documents, text transcription and named entity recognition are usually faced as separate subsequent tasks. This has the disadvantage that errors in the first module affect heavily the performance…

计算机视觉与模式识别 · 计算机科学 2018-03-23 Manuel Carbonell , Mauricio Villegas , Alicia Fornés , Josep Lladós

This work addresses the problem of Question Answering (QA) on handwritten document collections. Unlike typical QA and Visual Question Answering (VQA) formulations where the answer is a short text, we aim to locate a document snippet where…

计算机视觉与模式识别 · 计算机科学 2021-10-05 Minesh Mathew , Lluis Gomez , Dimosthenis Karatzas , CV Jawahar

This paper describes two approaches for content-based image retrieval and pattern spotting in document images using deep learning. The first approach uses a pre-trained CNN model to cope with the lack of training data, which is fine-tuned…

Handwritten document-image binarization is a semantic segmentation process to differentiate ink pixels from background pixels. It is one of the essential steps towards character recognition, writer identification, and script-style evolution…

计算机视觉与模式识别 · 计算机科学 2026-01-06 Maruf A. Dhali , Jan Willem de Wit , Lambert Schomaker

Visually Rich Document Understanding (VRDU) has emerged as a critical field in document intelligence, enabling automated extraction of key information from complex documents across domains such as medical, financial, and educational…

计算机视觉与模式识别 · 计算机科学 2025-06-03 Yihao Ding , Soyeon Caren Han , Yan Li , Josiah Poon

We present an overview of our triple extraction system for the ICDM 2019 Knowledge Graph Contest. Our system uses a pipeline-based approach to extract a set of triples from a given document. It offers a simple and effective solution to the…

计算与语言 · 计算机科学 2019-09-05 Michael Stewart , Majigsuren Enkhsaikhan , Wei Liu

The order in which the trajectory is executed is a powerful source of information for recognizers. However, there is still no general approach for recovering the trajectory of complex and long handwriting from static images. Complex…

计算机视觉与模式识别 · 计算机科学 2024-06-06 Moises Diaz , Gioele Crispo , Antonio Parziale , Angelo Marcelli , Miguel A. Ferrer

The Odeuropa Challenge on Olfactory Object Recognition aims to foster the development of object detection in the visual arts and to promote an olfactory perspective on digital heritage. Object detection in historical artworks is…

计算机视觉与模式识别 · 计算机科学 2023-01-25 Mathias Zinnen , Prathmesh Madhu , Ronak Kosti , Peter Bell , Andreas Maier , Vincent Christlein

Binarization of document images is an important pre-processing step in the field of document analysis. Traditional image binarization techniques usually rely on histograms or local statistics to identify a valid threshold to differentiate…

计算机视觉与模式识别 · 计算机科学 2024-01-23 Richin Sukesh , Mathias Seuret , Anguelos Nicolaou , Martin Mayr , Vincent Christlein

The indexing and searching of historical documents have garnered attention in recent years due to massive digitization efforts of important collections worldwide. Pure textual search in these corpora is a problem since optical character…

信息检索 · 计算机科学 2020-04-23 Taivanbat Badamdorj , Adiel Ben-Shalom , Nachum Dershowitz , Lior Wolf

This paper presents our solution for the ICDAR 2021 Competition on Scientific Table Image Recognition to LaTeX. This competition has two sub-tasks: Table Structure Reconstruction (TSR) and Table Content Reconstruction (TCR). We treat both…

计算机视觉与模式识别 · 计算机科学 2021-05-06 Yelin He , Xianbiao Qi , Jiaquan Ye , Peng Gao , Yihao Chen , Bingcong Li , Xin Tang , Rong Xiao

In this paper, we introduce iART: an open Web platform for art-historical research that facilitates the process of comparative vision. The system integrates various machine learning techniques for keyword- and content-based image retrieval…

The present paper introduces a group activity involving writing summaries of conference proceedings by volunteer participants. The rapid increase in scientific papers is a heavy burden for researchers, especially non-native speakers, who…

计算机视觉与模式识别 · 计算机科学 2022-03-18 Shintaro Yamamoto , Hirokatsu Kataoka , Ryota Suzuki , Seitaro Shinagawa , Shigeo Morishima

We find that large language models (LLMs) are more likely to modify human-written text than AI-generated text when tasked with rewriting. This tendency arises because LLMs often perceive AI-generated text as high-quality, leading to fewer…

计算与语言 · 计算机科学 2024-04-16 Chengzhi Mao , Carl Vondrick , Hao Wang , Junfeng Yang

Historical documents present many challenges for offline handwriting recognition systems, among them, the segmentation and labeling steps. Carefully annotated textlines are needed to train an HTR system. In some scenarios, transcripts are…

计算机视觉与模式识别 · 计算机科学 2018-11-20 Edgard Chammas , Chafic Mokbel , Laurence Likforman-Sulem

Despite the advancements in search engine features, ranking methods, technologies, and the availability of programmable APIs, current-day open-access digital libraries still rely on crawl-based approaches for acquiring their underlying…

信息检索 · 计算机科学 2016-04-19 Sujatha Das Gollapalli , Krutarth Patel , Cornelia Caragea

A lot of claims are made in social media posts, which may contain misinformation or fake news. Hence, it is crucial to identify claims as a first step towards claim verification. Given the huge number of social media posts, the task of…

计算与语言 · 计算机科学 2024-12-02 Soham Poddar , Biswajit Paul , Moumita Basu , Saptarshi Ghosh

Much computer vision research has focused on natural images, but technical documents typically consist of abstract images, such as charts, drawings, diagrams, and schematics. How well do general web search engines discover abstract images?…

计算机视觉与模式识别 · 计算机科学 2022-11-07 Shawn M. Jones , Diane Oyen

The aim of a Content-Based Image Retrieval (CBIR) system, also known as Query by Image Content (QBIC), is to help users to retrieve relevant images based on their contents. CBIR technologies provide a method to find images in large…

计算机视觉与模式识别 · 计算机科学 2020-07-10 Aman Chadha , Sushmit Mallik , Ravdeep Johar

Motion blur is one of the most common degradation artifacts in dynamic scene photography. This paper reviews the NTIRE 2020 Challenge on Image and Video Deblurring. In this challenge, we present the evaluation results from 3 competition…

计算机视觉与模式识别 · 计算机科学 2020-05-12 Seungjun Nah , Sanghyun Son , Radu Timofte , Kyoung Mu Lee