中文
相关论文

相关论文: ICDAR 2023 Competition on Reading the Seal Title

200 篇论文

We organize a competition on hierarchical text detection and recognition. The competition is aimed to promote research into deep learning models and systems that can jointly perform text detection and recognition and geometric layout…

计算机视觉与模式识别 · 计算机科学 2023-05-18 Shangbang Long , Siyang Qin , Dmitry Panteleev , Alessandro Bissacco , Yasuhisa Fujii , Michalis Raptis

With hundreds of thousands of electronic chip components are being manufactured every day, chip manufacturers have seen an increasing demand in seeking a more efficient and effective way of inspecting the quality of printed texts on chip…

计算机视觉与模式识别 · 计算机科学 2021-07-13 Chun Chet Ng , Akmalul Khairi Bin Nazaruddin , Yeong Khang Lee , Xinyu Wang , Yuliang Liu , Chee Seng Chan , Lianwen Jin , Yipeng Sun , Lixin Fan

Chinese scene text reading is one of the most challenging problems in computer vision and has attracted great interest. Different from English text, Chinese has more than 6000 commonly used characters and Chinesecharacters can be arranged…

Structured text extraction is one of the most valuable and challenging application directions in the field of Document AI. However, the scenarios of past benchmarks are limited, and the corresponding evaluation protocols usually focus on…

With the growing cosmopolitan culture of modern cities, the need of robust Multi-Lingual scene Text (MLT) detection and recognition systems has never been more immense. With the goal to systematically benchmark and push the state-of-the-art…

Robust text reading from street view images provides valuable information for various applications. Performance improvement of existing methods in such a challenging scenario heavily relies on the amount of fully annotated training data,…

计算机视觉与模式识别 · 计算机科学 2019-09-18 Yipeng Sun , Zihan Ni , Chee-Kheng Chng , Yuliang Liu , Canjie Luo , Chun Chet Ng , Junyu Han , Errui Ding , Jingtuo Liu , Dimosthenis Karatzas , Chee Seng Chan , Lianwen Jin

Recently, video text detection, tracking, and recognition in natural scenes are becoming very popular in the computer vision community. However, most existing algorithms and benchmarks focus on common text cases (e.g., normal size, density)…

计算机视觉与模式识别 · 计算机科学 2023-04-11 Weijia Wu , Yuzhong Zhao , Zhuang Li , Jiahong Li , Mike Zheng Shou , Umapada Pal , Dimosthenis Karatzas , Xiang Bai

Transforming documents into machine-processable representations is a challenging task due to their complex structures and variability in formats. Recovering the layout structure and content from PDF files or scanned material has remained a…

计算机视觉与模式识别 · 计算机科学 2023-08-22 Christoph Auer , Ahmed Nassar , Maksym Lysak , Michele Dolfi , Nikolaos Livathinos , Peter Staar

Scanned receipts OCR and key information extraction (SROIE) represent the processeses of recognizing text from scanned receipts and extracting key texts from them and save the extracted tests to structured documents. SROIE plays critical…

人工智能 · 计算机科学 2021-03-19 Zheng Huang , Kai Chen , Jianhua He , Xiang Bai , Dimosthenis Karatzas , Shjian Lu , C. V. Jawahar

Tables present important information concisely in many scientific documents. Visual features like mathematical symbols, equations, and spanning cells make structure and content extraction from tables embedded in research documents…

信息检索 · 计算机科学 2021-11-12 Pratik Kayal , Mrinal Anand , Harsh Desai , Mayank Singh

This paper presents our solution for ICDAR 2021 competition on scientific literature parsing taskB: table recognition to HTML. In our method, we divide the table content recognition task into foursub-tasks: table structure recognition, text…

计算机视觉与模式识别 · 计算机科学 2021-05-06 Jiaquan Ye , Xianbiao Qi , Yelin He , Yihao Chen , Dengyi Gu , Peng Gao , Rong Xiao

This paper reports the ICDAR2019 Robust Reading Challenge on Arbitrary-Shaped Text (RRC-ArT) that consists of three major challenges: i) scene text detection, ii) scene text recognition, and iii) scene text spotting. A total of 78…

Document Image Machine Translation (DIMT) seeks to translate text embedded in document images from one language to another by jointly modeling both textual content and page layout, bridging optical character recognition (OCR) and natural…

计算机视觉与模式识别 · 计算机科学 2026-03-11 Yaping Zhang , Yupu Liang , Zhiyang Zhang , Zhiyuan Chen , Lu Xiang , Yang Zhao , Yu Zhou , Chengqing Zong

Scene video text spotting (SVTS) is a very important research topic because of many real-life applications. However, only a little effort has put to spotting scene video text, in contrast to massive studies of scene text spotting in static…

计算机视觉与模式识别 · 计算机科学 2021-07-27 Zhanzhan Cheng , Jing Lu , Baorui Zou , Shuigeng Zhou , Fei Wu

Text image super-resolution is a challenging yet open research problem in the computer vision community. In particular, low-resolution images hamper the performance of typical optical character recognition (OCR) systems. In this article, we…

计算机视觉与模式识别 · 计算机科学 2015-06-09 Chao Dong , Ximei Zhu , Yubin Deng , Chen Change Loy , Yu Qiao

Scientific literature contain important information related to cutting-edge innovations in diverse domains. Advances in natural language processing have been driving the fast development in automated information extraction from scientific…

信息检索 · 计算机科学 2021-06-29 Antonio Jimeno Yepes , Xu Zhong , Douglas Burdick

Recently, text detection and recognition in natural scenes are becoming increasing popular in the computer vision community as well as the document analysis community. However, majority of the existing ideas, algorithms and systems are…

计算机视觉与模式识别 · 计算机科学 2015-06-11 Xinyu Zhou , Shuchang Zhou , Cong Yao , Zhimin Cao , Qi Yin

This paper describes the experimental framework and results of the ICDAR 2021 Competition on On-Line Signature Verification (SVC 2021). The goal of SVC 2021 is to evaluate the limits of on-line signature verification systems on popular…

This paper presents our solution for the ICDAR 2021 Competition on Scientific Table Image Recognition to LaTeX. This competition has two sub-tasks: Table Structure Reconstruction (TSR) and Table Content Reconstruction (TCR). We treat both…

计算机视觉与模式识别 · 计算机科学 2021-05-06 Yelin He , Xianbiao Qi , Jiaquan Ye , Peng Gao , Yihao Chen , Bingcong Li , Xin Tang , Rong Xiao

This paper presents final results of ICDAR 2019 Scene Text Visual Question Answering competition (ST-VQA). ST-VQA introduces an important aspect that is not addressed by any Visual Question Answering system up to date, namely the…

计算机视觉与模式识别 · 计算机科学 2019-07-02 Ali Furkan Biten , Rubèn Tito , Andres Mafla , Lluis Gomez , Marçal Rusiñol , Minesh Mathew , C. V. Jawahar , Ernest Valveny , Dimosthenis Karatzas
‹ 上一页 1 2 3 10 下一页 ›