中文
相关论文

相关论文: Sliding window approach based Text Binarisation fr…

200 篇论文

Binarization is a popular first step towards text extraction in historical artifacts. Stone inscription images pose severe challenges for binarization due to poor contrast between etched characters and the stone background, non-uniform…

计算机视觉与模式识别 · 计算机科学 2026-01-08 Pratyush Jena , Amal Joseph , Arnav Sharma , Ravi Kiran Sarvadevabhatla

Handwritten document-image binarization is a semantic segmentation process to differentiate ink pixels from background pixels. It is one of the essential steps towards character recognition, writer identification, and script-style evolution…

计算机视觉与模式识别 · 计算机科学 2026-01-06 Maruf A. Dhali , Jan Willem de Wit , Lambert Schomaker

Bimodal objects, such as the checkerboard pattern used in camera calibration, markers for object tracking, and text on road signs, to name a few, are prevalent in our daily lives and serve as a visual form to embed information that can be…

计算机视觉与模式识别 · 计算机科学 2024-02-21 Shijie Lin , Xiang Zhang , Lei Yang , Lei Yu , Bin Zhou , Xiaowei Luo , Wenping Wang , Jia Pan

In recent years, recognition of text from natural scene image and video frame has got increased attention among the researchers due to its various complexities and challenges. Because of low resolution, blurring effect, complex background,…

计算机视觉与模式识别 · 计算机科学 2017-07-31 Ayan Kumar Bhunia , Gautam Kumar , Partha Pratim Roy , R. Balasubramanian , Umapada Pal

The existing Optical Character Recognition (OCR) systems are capable of recognizing images with horizontal texts. However, when the rotation of the texts increases, it becomes harder to recognizing these texts. The performance of the OCR…

计算机视觉与模式识别 · 计算机科学 2022-09-20 Michael Yang , Yuan Lin , ChiuMan Ho

Recently, segmentation-based scene text detection methods have drawn extensive attention in the scene text detection field, because of their superiority in detecting the text instances of arbitrary shapes and extreme aspect ratios,…

计算机视觉与模式识别 · 计算机科学 2022-02-22 Minghui Liao , Zhisheng Zou , Zhaoyi Wan , Cong Yao , Xiang Bai

To recognize textures many methods have been developed along the years. However, texture datasets may be hard to be classified due to artefacts such as a variety of scale, illumination and noise. This paper proposes the application of…

计算机视觉与模式识别 · 计算机科学 2016-12-21 Mariane Barros Neiva , Antoine Manzanera , Odemir Martinez Bruno

In this paper, we propose a gradient difference based approach to text localization in videos and scene images. The input video frame/ image is first compressed using multilevel 2-D wavelet transform. The edge information of the…

计算机视觉与模式识别 · 计算机科学 2015-02-23 B. H. Shekar , Smitha M. L

Thresholding converts a greyscale image into a binary image, and is thus often a necessary segmentation step in image processing. For a human viewer however, thresholding usually has a negative impact on the legibility of document images.…

计算机视觉与模式识别 · 计算机科学 2023-01-20 Christoph Dalitz

A novel method to convert color/multi-spectral images to gray-level images is introduced to increase the performance of document binarization methods. The method uses the distribution of the pixel data of the input document image in a color…

计算机视觉与模式识别 · 计算机科学 2013-06-27 Reza Farrahi Moghaddam , Shaohua Chen , Rachid Hedjam , Mohamed Cheriet

Business card images are of multiple natures as these often contain graphics, pictures and texts of various fonts and sizes both in background and foreground. So, the conventional binarization techniques designed for document images can not…

计算机视觉与模式识别 · 计算机科学 2010-03-09 Ayatullah Faruk Mollah , Subhadip Basu , Nibaran Das , Ram Sarkar , Mita Nasipuri , Mahantapas Kundu

In recent years, partial differential equation (PDE) systems have been successfully applied to the binarization of text images, achieving promising results. Inspired by the DH model and incorporating a novel image modeling approach, this…

计算机视觉与模式识别 · 计算机科学 2025-03-12 Youjin Liu , Yu Wang

Digital camera and mobile document image acquisition are new trends arising in the world of Optical Character Recognition and text detection. In some cases, such process integrates many distortions and produces poorly scanned text or…

计算机视觉与模式识别 · 计算机科学 2015-09-14 Abdeslam El Harraj , Naoufal Raissouni

Extraction and recognition of Bangla text from video frame images is challenging due to complex color background, low-resolution etc. In this paper, we propose an algorithm for extraction and recognition of Bangla text form such video…

计算机视觉与模式识别 · 计算机科学 2014-01-07 Souvik Bhowmick , Purnendu Banerjee

Sliding window approaches have been widely used for object recognition tasks in recent years. They guarantee an investigation of the entire input image for the object to be detected and allow a localization of that object. Despite the…

计算机视觉与模式识别 · 计算机科学 2018-08-07 Julian Müller , Andreas Fregin , Klaus Dietmayer

Recently, segmentation-based methods are quite popular in scene text detection, as the segmentation results can more accurately describe scene text of various shapes such as curve text. However, the post-processing of binarization is…

计算机视觉与模式识别 · 计算机科学 2019-12-04 Minghui Liao , Zhaoyi Wan , Cong Yao , Kai Chen , Xiang Bai

The main aim of this paper is to propose a color texture classification approach which uses color sensor information and texture features jointly. High accuracy, low noise sensitivity and low computational complexity are specified aims for…

计算机视觉与模式识别 · 计算机科学 2019-06-27 Shervan Fekri-Ershad , Farshad Tajeripour

Foreground-background separation is an important problem in document image analysis. Popular unsupervised binarization methods (such as the Sauvola's algorithm) employ adaptive thresholding to classify pixels as foreground or background. In…

计算机视觉与模式识别 · 计算机科学 2022-04-11 Soumyadeep Dey , Pratik Jawanpuria

Contrast enhancement is an important area of research for the image analysis. Over the decade, the researcher worked on this domain to develop an efficient and adequate algorithm. The proposed method will enhance the contrast of image using…

多媒体 · 计算机科学 2012-05-08 Aroop Mukherjee , Soumen Kanrar

Document comparison typically relies on optical character recognition (OCR) as its core technology. However, OCR requires the selection of appropriate language models for each document and the performance of multilingual or hybrid models…

计算机视觉与模式识别 · 计算机科学 2024-12-06 Doyoung Park , Naresh Reddy Yarram , Sunjin Kim , Minkyu Kim , Seongho Cho , Taehee Lee