English
Related papers

Related papers: LS-HDIB: A Large Scale Handwritten Document Image …

200 papers

For many fundamental scene understanding tasks, it is difficult or impossible to obtain per-pixel ground truth labels from real images. We address this challenge by introducing Hypersim, a photorealistic synthetic dataset for holistic…

Computer Vision and Pattern Recognition · Computer Science 2021-08-19 Mike Roberts , Jason Ramapuram , Anurag Ranjan , Atulit Kumar , Miguel Angel Bautista , Nathan Paczan , Russ Webb , Joshua M. Susskind

This paper tackles two key challenges: detecting small, dense, and overlapping objects (a major hurdle in computer vision) and improving the quality of noisy images, especially those encountered in industrial environments. [1, 2]. Our focus…

Computer Vision and Pattern Recognition · Computer Science 2025-09-04 Oussama Messai , Abbass Zein-Eddine , Abdelouahid Bentamou , Mickaël Picq , Nicolas Duquesne , Stéphane Puydarrieux , Yann Gavet

We introduce XYZ-IBD, a bin-picking dataset for 6D pose estimation that captures real-world industrial complexity, including challenging object geometries, reflective materials, severe occlusions, and dense clutter. The dataset reflects…

Computer Vision and Pattern Recognition · Computer Science 2025-06-17 Junwen Huang , Jizhong Liang , Jiaqi Hu , Martin Sundermeyer , Peter KT Yu , Nassir Navab , Benjamin Busam

Underwater image restoration is of significant importance in unveiling the underwater world. Numerous techniques and algorithms have been developed in the past decades. However, due to fundamental difficulties associated with…

Image and Video Processing · Electrical Eng. & Systems 2021-06-22 Junlin Han , Mehrdad Shoeiby , Tim Malthus , Elizabeth Botha , Janet Anstee , Saeed Anwar , Ran Wei , Mohammad Ali Armin , Hongdong Li , Lars Petersson

The state of the art in human-centric computer vision achieves high accuracy and robustness across a diverse range of tasks. The most effective models in this domain have billions of parameters, thus requiring extremely large datasets,…

Computer Vision and Pattern Recognition · Computer Science 2025-07-22 Fatemeh Saleh , Sadegh Aliakbarian , Charlie Hewitt , Lohit Petikam , Xiao-Xian , Antonio Criminisi , Thomas J. Cashman , Tadas Baltrušaitis

Handwritten word recognition from document images using deep learning is an active research area in the field of Document Image Analysis and Recognition. In the present era of Big data, since more and more documents are being generated and…

Computer Vision and Pattern Recognition · Computer Science 2023-02-20 Bulla Rajesh , Abhishek Kumar Gupta , Ayush Raj , Mohammed Javed , Shiv Ram Dubey

There is a need for information retrieval from large collections of low-resolution (LR) binary document images, which can be found in digital libraries across the world, where the high-resolution (HR) counterpart is not available. This…

Computer Vision and Pattern Recognition · Computer Science 2018-12-07 Ram Krishna Pandey , K Vignesh , A G Ramakrishnan , Chandrahasa B

The advent of a panoply of resource limited devices opens up new challenges in the design of computer vision algorithms with a clear compromise between accuracy and computational requirements. In this paper we present new binary image…

Computer Vision and Pattern Recognition · Computer Science 2021-08-20 Iago Suárez , José M. Buenaposada , Luis Baumela

High Dynamic Range (HDR) content (i.e., images and videos) has a broad range of applications. However, capturing HDR content from real-world scenes is expensive and time-consuming. Therefore, the challenging task of reconstructing visually…

Computer Vision and Pattern Recognition · Computer Science 2024-03-28 Hrishav Bakul Barua , Kalin Stefanov , KokSheik Wong , Abhinav Dhall , Ganesh Krishnasamy

In this study, we evaluated four binarization methods. Locality-Sensitive Hashing (LSH), Iterative Quantization (ITQ), Kernel-based Supervised Hashing (KSH), and Supervised Discrete Hashing (SDH) on the ODIR dataset using deep feature…

Image and Video Processing · Electrical Eng. & Systems 2026-01-09 Nedim Muzoglu

Capturing the shape and spatially-varying appearance (SVBRDF) of an object from images is a challenging task that has applications in both computer vision and graphics. Traditional optimization-based approaches often need a large number of…

Computer Vision and Pattern Recognition · Computer Science 2021-05-20 Mark Boss , Varun Jampani , Kihwan Kim , Hendrik P. A. Lensch , Jan Kautz

The advent of deep learning has brought a revolutionary transformation to image denoising techniques. However, the persistent challenge of acquiring noise-clean pairs for supervised methods in real-world scenarios remains formidable,…

Image and Video Processing · Electrical Eng. & Systems 2024-03-26 Dan Zhang , Fangfang Zhou , Felix Albu , Yuanzhou Wei , Xiao Yang , Yuan Gu , Qiang Li

There are two types of information in each handwritten word image: explicit information which can be easily read or derived directly, such as lexical content or word length, and implicit attributes such as the author's identity. Whether…

Computer Vision and Pattern Recognition · Computer Science 2018-11-20 Sheng He , Lambert Schomaker

Image denoising is often empowered by accurate prior information. In recent years, data-driven neural network priors have shown promising performance for RGB natural image denoising. Compared to classic handcrafted priors (e.g., sparsity…

Image and Video Processing · Electrical Eng. & Systems 2022-02-16 Yu-Chun Miao , Xi-Le Zhao , Xiao Fu , Jian-Li Wang , Yu-Bang Zheng

Self-supervised learning (SSL) methods targeting scene images have seen a rapid growth recently, and they mostly rely on either a dedicated dense matching mechanism or a costly unsupervised object discovery module. This paper shows that…

Computer Vision and Pattern Recognition · Computer Science 2023-10-02 Ke Zhu , Minghao Fu , Jianxin Wu

We consider the challenging problem of predicting intrinsic object properties from a single image by exploiting differentiable renderers. Many previous learning-based approaches for inverse graphics adopt rasterization-based renderers and…

Computer Vision and Pattern Recognition · Computer Science 2021-11-02 Wenzheng Chen , Joey Litalien , Jun Gao , Zian Wang , Clement Fuji Tsang , Sameh Khamis , Or Litany , Sanja Fidler

Visual text rendering, which aims to accurately integrate specified textual content within generated images, is critical for various applications such as commercial design. Despite recent advances, current methods struggle with long-tail…

Computer Vision and Pattern Recognition · Computer Science 2025-05-13 Shuhan Zhuang , Mengqi Huang , Fengyi Fu , Nan Chen , Bohan Lei , Zhendong Mao

The classification of imbalanced data streams, which have unequal class distributions, is a key difficulty in machine learning, especially when dealing with multiple classes. While binary imbalanced data stream classification tasks have…

Machine Learning · Computer Science 2025-06-26 Soheil Abadifard , Fazli Can

One of the challenges of handwriting recognition is to transcribe a large number of vastly different writing styles. State-of-the-art approaches do not explicitly use information about the writer's style, which may be limiting overall…

Computer Vision and Pattern Recognition · Computer Science 2025-05-01 Jan Kohút , Michal Hradiš , Martin Kišš

The rapid growth of image data has led to the development of advanced image processing and computer vision techniques, which are crucial in various applications such as image classification, image segmentation, and pattern recognition.…

Computer Vision and Pattern Recognition · Computer Science 2024-07-29 Zeinab Sedaghatjoo , Hossein Hosseinzadeh , Bahram Sadeghi Bigham