English
Related papers

Related papers: ICDAR 2023 Competition on Structured Text Extracti…

200 papers

Transforming documents into machine-processable representations is a challenging task due to their complex structures and variability in formats. Recovering the layout structure and content from PDF files or scanned material has remained a…

Computer Vision and Pattern Recognition · Computer Science 2023-08-22 Christoph Auer , Ahmed Nassar , Maksym Lysak , Michele Dolfi , Nikolaos Livathinos , Peter Staar

Scene video text spotting (SVTS) is a very important research topic because of many real-life applications. However, only a little effort has put to spotting scene video text, in contrast to massive studies of scene text spotting in static…

Computer Vision and Pattern Recognition · Computer Science 2021-07-27 Zhanzhan Cheng , Jing Lu , Baorui Zou , Shuigeng Zhou , Fei Wu

We organize a competition on hierarchical text detection and recognition. The competition is aimed to promote research into deep learning models and systems that can jointly perform text detection and recognition and geometric layout…

Computer Vision and Pattern Recognition · Computer Science 2023-05-18 Shangbang Long , Siyang Qin , Dmitry Panteleev , Alessandro Bissacco , Yasuhisa Fujii , Michalis Raptis

Recently, video text detection, tracking, and recognition in natural scenes are becoming very popular in the computer vision community. However, most existing algorithms and benchmarks focus on common text cases (e.g., normal size, density)…

Computer Vision and Pattern Recognition · Computer Science 2023-04-11 Weijia Wu , Yuzhong Zhao , Zhuang Li , Jiahong Li , Mike Zheng Shou , Umapada Pal , Dimosthenis Karatzas , Xiang Bai

Robust text reading from street view images provides valuable information for various applications. Performance improvement of existing methods in such a challenging scenario heavily relies on the amount of fully annotated training data,…

Computer Vision and Pattern Recognition · Computer Science 2019-09-18 Yipeng Sun , Zihan Ni , Chee-Kheng Chng , Yuliang Liu , Canjie Luo , Chun Chet Ng , Junyu Han , Errui Ding , Jingtuo Liu , Dimosthenis Karatzas , Chee Seng Chan , Lianwen Jin

Document Image Machine Translation (DIMT) seeks to translate text embedded in document images from one language to another by jointly modeling both textual content and page layout, bridging optical character recognition (OCR) and natural…

Computer Vision and Pattern Recognition · Computer Science 2026-03-11 Yaping Zhang , Yupu Liang , Zhiyang Zhang , Zhiyuan Chen , Lu Xiang , Yang Zhao , Yu Zhou , Chengqing Zong

Text line segmentation is a critical step in handwritten document image analysis. Segmenting text lines in historical handwritten documents, however, presents unique challenges due to irregular handwriting, faded ink, and complex layouts…

Computer Vision and Pattern Recognition · Computer Science 2025-09-17 Silvia Zottin , Axel De Nardin , Giuseppe Branca , Claudio Piciarelli , Gian Luca Foresti

Visually Rich Document Understanding (VRDU) has emerged as a critical field in document intelligence, enabling automated extraction of key information from complex documents across domains such as medical, financial, and educational…

Computer Vision and Pattern Recognition · Computer Science 2025-06-03 Yihao Ding , Soyeon Caren Han , Yan Li , Josiah Poon

Reading seal title text is a challenging task due to the variable shapes of seals, curved text, background noise, and overlapped text. However, this important element is commonly found in official and financial scenarios, and has not…

Computer Vision and Pattern Recognition · Computer Science 2023-06-07 Wenwen Yu , Mingyu Liu , Mingrui Chen , Ning Lu , Yinlong Wen , Yuliang Liu , Dimosthenis Karatzas , Xiang Bai

Tables present important information concisely in many scientific documents. Visual features like mathematical symbols, equations, and spanning cells make structure and content extraction from tables embedded in research documents…

Information Retrieval · Computer Science 2021-11-12 Pratik Kayal , Mrinal Anand , Harsh Desai , Mayank Singh

Understanding visually-rich business documents to extract structured data and automate business workflows has been receiving attention both in academia and industry. Although recent multi-modal language models have achieved impressive…

Computation and Language · Computer Science 2023-09-19 Zilong Wang , Yichao Zhou , Wei Wei , Chen-Yu Lee , Sandeep Tata

This paper describes the short-term competition on the Components Segmentation Task of Document Photos that was prepared in the context of the 16th International Conference on Document Analysis and Recognition (ICDAR 2021). This competition…

Computer Vision and Pattern Recognition · Computer Science 2021-07-12 Celso A. M. Lopes Junior , Ricardo B. das Neves Junior , Byron L. D. Bezerra , Alejandro H. Toselli , Donato Impedovo

This paper presents final results of ICDAR 2019 Scene Text Visual Question Answering competition (ST-VQA). ST-VQA introduces an important aspect that is not addressed by any Visual Question Answering system up to date, namely the…

Computer Vision and Pattern Recognition · Computer Science 2019-07-02 Ali Furkan Biten , Rubèn Tito , Andres Mafla , Lluis Gomez , Marçal Rusiñol , Minesh Mathew , C. V. Jawahar , Ernest Valveny , Dimosthenis Karatzas

Scientific literature contain important information related to cutting-edge innovations in diverse domains. Advances in natural language processing have been driving the fast development in automated information extraction from scientific…

Information Retrieval · Computer Science 2021-06-29 Antonio Jimeno Yepes , Xu Zhong , Douglas Burdick

With hundreds of thousands of electronic chip components are being manufactured every day, chip manufacturers have seen an increasing demand in seeking a more efficient and effective way of inspecting the quality of printed texts on chip…

Computer Vision and Pattern Recognition · Computer Science 2021-07-13 Chun Chet Ng , Akmalul Khairi Bin Nazaruddin , Yeong Khang Lee , Xinyu Wang , Yuliang Liu , Chee Seng Chan , Lianwen Jin , Yipeng Sun , Lixin Fan

Recently, text detection and recognition in natural scenes are becoming increasing popular in the computer vision community as well as the document analysis community. However, majority of the existing ideas, algorithms and systems are…

Computer Vision and Pattern Recognition · Computer Science 2015-06-11 Xinyu Zhou , Shuchang Zhou , Cong Yao , Zhimin Cao , Qi Yin

This paper presents our solution for ICDAR 2021 competition on scientific literature parsing taskB: table recognition to HTML. In our method, we divide the table content recognition task into foursub-tasks: table structure recognition, text…

Computer Vision and Pattern Recognition · Computer Science 2021-05-06 Jiaquan Ye , Xianbiao Qi , Yelin He , Yihao Chen , Dengyi Gu , Peng Gao , Rong Xiao

Text line detection is crucial for any application associated with Automatic Text Recognition or Keyword Spotting. Modern algorithms perform good on well-established datasets since they either comprise clean data or simple/homogeneous page…

Computer Vision and Pattern Recognition · Computer Science 2017-12-12 Tobias Grüning , Roger Labahn , Markus Diem , Florian Kleber , Stefan Fiel

The lack of data for information extraction (IE) from semi-structured business documents is a real problem for the IE community. Publications relying on large-scale datasets use only proprietary, unpublished data due to the sensitive nature…

Machine Learning · Computer Science 2023-01-31 Štěpán Šimsa , Milan Šulc , Matyáš Skalický , Yash Patel , Ahmed Hamdi

Scanned receipts OCR and key information extraction (SROIE) represent the processeses of recognizing text from scanned receipts and extracting key texts from them and save the extracted tests to structured documents. SROIE plays critical…

Artificial Intelligence · Computer Science 2021-03-19 Zheng Huang , Kai Chen , Jianhua He , Xiang Bai , Dimosthenis Karatzas , Shjian Lu , C. V. Jawahar
‹ Prev 1 2 3 10 Next ›