中文
相关论文

相关论文: Optical Braille Recognition using Circular Hough T…

200 篇论文

In today's world, time is a very important resource. In our busy lives, most of us hardly have time to read the complete news so what we have to do is just go through the headlines and satisfy ourselves with that. As a result, we might miss…

人机交互 · 计算机科学 2020-01-06 Mona teja K , Mohan Sai. S , H S S S Raviteja D , Sai Kushagra P

Blur is an image degradation that is difficult to remove. Invariants with respect to blur offer an alternative way of a~description and recognition of blurred images without any deblurring. In this paper, we present an original unified…

计算机视觉与模式识别 · 计算机科学 2023-08-03 Jan Flusser , Matej Lebl , Matteo Pedone , Filip Sroubek , Jitka Kostkova

Thresholding converts a greyscale image into a binary image, and is thus often a necessary segmentation step in image processing. For a human viewer however, thresholding usually has a negative impact on the legibility of document images.…

计算机视觉与模式识别 · 计算机科学 2023-01-20 Christoph Dalitz

In blind motion deblurring, leading methods today tend towards highly non-convex approximations of the l0-norm, especially in the image regularization term. In this paper, we propose a simple, effective and fast approach for the estimation…

计算机视觉与模式识别 · 计算机科学 2015-01-23 Wen-Ze Shao , Hai-Bo Li , Michael Elad

Anomaly detection is the process of identifying atypical data samples that significantly deviate from the majority of the dataset. In the realm of clinical screening and diagnosis, detecting abnormalities in medical images holds great…

计算机视觉与模式识别 · 计算机科学 2023-10-11 Xianyao Hu , Congming Jin

When a reader encounters a word in English, they split the word into smaller orthographic units in the process of recognizing its meaning. For example, "rough", when split according to phonemes, is decomposed as r-ou-gh (not as r-o-ugh or…

人机交互 · 计算机科学 2025-08-26 Matthew Termuende , Kevin Larson , Miguel Nacenta

Historical documents frequently suffer from damage and inconsistencies, including missing or illegible text resulting from issues such as holes, ink problems, and storage damage. These missing portions or gaps are referred to as lacunae. In…

计算机视觉与模式识别 · 计算机科学 2024-07-02 Jaydeep Borkar , David A. Smith

Recovering sharp images from dual-pixel (DP) pairs with disparity-dependent blur is a challenging task.~Existing blur map-based deblurring methods have demonstrated promising results. In this paper, we propose, to the best of our knowledge,…

计算机视觉与模式识别 · 计算机科学 2023-11-22 Hao Yang , Liyuan Pan , Yan Yang , Richard Hartley , Miaomiao Liu

A comprehensive assessment of retinal health demands reliable and precise methods to measure localized blood perfusion. Despite considerable advancements in imaging techniques, such as indocyanine green and fluorescein angiography, along…

医学物理 · 物理学 2024-03-15 Michael Atlan

We address the challenging problem of Natural Language Comprehension beyond plain-text documents by introducing the TILT neural network architecture which simultaneously learns layout information, visual features, and textual semantics.…

计算与语言 · 计算机科学 2021-07-13 Rafał Powalski , Łukasz Borchmann , Dawid Jurkiewicz , Tomasz Dwojak , Michał Pietruszka , Gabriela Pałka

An increasing number of Chinese people are troubled by different degrees of visual impairment, which has made the modal conversion between a single image or video frame in the visual field and the audio expressing the same information a…

声音 · 计算机科学 2024-07-22 Chun Xu , En-Wei Sun

Navigation and obstacle avoidance are some of the hardest tasks for the visually impaired. Recent research projects have proposed technological solutions to tackle this problem. So far most systems fail to provide multidimensional feedback…

人机交互 · 计算机科学 2022-01-13 Manuel Zahn , Armaghan Ahmad Khan

Text binarisation process classifies individual pixels as text or background in the textual images. Binarization is necessary to bridge the gap between localization and recognition by OCR. This paper presents Sliding window method to…

计算机视觉与模式识别 · 计算机科学 2010-03-19 Chitrakala Gopalan , D. Manjula

Document clustering is an unsupervised approach in which a large collection of documents (corpus) is subdivided into smaller, meaningful, identifiable, and verifiable sub-groups (clusters). Meaningful representation of documents and…

信息检索 · 计算机科学 2014-12-08 Muhammad Rafi , Farnaz Amin , Mohammad Shahid Shaikh

In this work, we propose a new framework, called Document Image Transformer (DocTr), to address the issue of geometry and illumination distortion of the document images. Specifically, DocTr consists of a geometric unwarping transformer and…

计算机视觉与模式识别 · 计算机科学 2022-10-11 Hao Feng , Yuechen Wang , Wengang Zhou , Jiajun Deng , Houqiang Li

This study aims to conduct an extensive detailed analysis of the Odia Braille reading comprehension among students with visual disability. Specifically, the study explores their reading speed and hand or finger movements. The study also…

计算与语言 · 计算机科学 2023-10-13 Monnie Parida , Manjira Sinha , Anupam Basu , Pabitra Mitra

In this paper, we suggest a new neural network architecture for vanishing point detection in images. The key element is the use of the direct and transposed Fast Hough Transforms separated by convolutional layer blocks with standard…

计算机视觉与模式识别 · 计算机科学 2020-11-11 A. Sheshkus , A. Chirvonaya , D. Matveev , D. Nikolaev , V. L. Arlazarov

To recognize textures many methods have been developed along the years. However, texture datasets may be hard to be classified due to artefacts such as a variety of scale, illumination and noise. This paper proposes the application of…

计算机视觉与模式识别 · 计算机科学 2016-12-21 Mariane Barros Neiva , Antoine Manzanera , Odemir Martinez Bruno

In the real-life environments, due to the sudden appearance of windows, lights, and objects blocking the light source, the visual SLAM system can easily capture the low-contrast images caused by over-exposure or over-darkness. At this time,…

机器人学 · 计算机科学 2019-02-12 Yinghong Fang , Guangcun Shan , Xin Li , Wenliang Liu , Tian Wang , Hichem Snoussi

Moving objects are frequently seen in daily life and usually appear blurred in images due to their motion. While general object retrieval is a widely explored area in computer vision, it primarily focuses on sharp and static objects, and…

计算机视觉与模式识别 · 计算机科学 2024-07-19 Rong Zou , Marc Pollefeys , Denys Rozumnyi