English
Related papers

Related papers: Single Document Image Highlight Removal via A Larg…

200 papers

Large ground-truth datasets and recent advances in deep learning techniques have been useful for layout detection. However, because of the restricted layout diversity of these datasets, training on them requires a sizable number of…

Computer Vision and Pattern Recognition · Computer Science 2024-04-22 Avinash Anand , Raj Jaiswal , Mohit Gupta , Siddhesh S Bangar , Pijush Bhuyan , Naman Lal , Rajeev Singh , Ritika Jha , Rajiv Ratn Shah , Shin'ichi Satoh

Extending CLIP models to semantic segmentation remains challenging due to the misalignment between their image-level pre-training objectives and the pixel-level visual understanding required for dense prediction. While prior efforts have…

Computer Vision and Pattern Recognition · Computer Science 2025-10-29 Jinxin Zhou , Jiachen Jiang , Zhihui Zhu

Merging multi-exposure images is a common approach for obtaining high dynamic range (HDR) images, with the primary challenge being the avoidance of ghosting artifacts in dynamic scenes. Recent methods have proposed using deep neural…

Computer Vision and Pattern Recognition · Computer Science 2024-02-29 Zhilu Zhang , Haoyu Wang , Shuai Liu , Xiaotao Wang , Lei Lei , Wangmeng Zuo

We present HetNet (Multi-level \textbf{Het}erogeneous \textbf{Net}work), a highly efficient mirror detection network. Current mirror detection methods focus more on performance than efficiency, limiting the real-time applications (such as…

Computer Vision and Pattern Recognition · Computer Science 2022-11-29 Ruozhen He , Jiaying Lin , Rynson W. H. Lau

This paper reviews the first challenge on high-dynamic range (HDR) imaging that was part of the New Trends in Image Restoration and Enhancement (NTIRE) workshop, held in conjunction with CVPR 2021. This manuscript focuses on the newly…

Computer Vision and Pattern Recognition · Computer Science 2021-06-04 Eduardo Pérez-Pellitero , Sibi Catley-Chandar , Aleš Leonardis , Radu Timofte

Single image deraining task is still a very challenging task due to its ill-posed nature in reality. Recently, researchers have tried to fix this issue by training the CNN-based end-to-end models, but they still cannot extract the negative…

Image and Video Processing · Electrical Eng. & Systems 2019-08-29 Yanyan Wei , Zhao Zhang , Haijun Zhang , Richang Hong , Meng Wang

Multimodal image registration is a fundamental task and a prerequisite for downstream cross-modal analysis. Despite recent progress in shared feature extraction and multi-scale architectures, two key limitations remain. First, some methods…

Computer Vision and Pattern Recognition · Computer Science 2026-03-23 Chunlei Zhang , Jiahao Xia , Yun Xiao , Bo Jiang , Jian Zhang

Document shadow removal is essential for enhancing the clarity of digitized documents. Preserving high-frequency details (e.g., text edges and lines) is critical in this process because shadows often obscure or distort fine structures. This…

Computer Vision and Pattern Recognition · Computer Science 2025-12-10 Chaewon Kim , Seoyeon Lee , Jonghyuk Park

Existing document-level machine translation resources are only available for a handful of languages, mostly high-resourced ones. To facilitate the training and evaluation of document-level translation and, more broadly, long-context…

Computation and Language · Computer Science 2025-10-01 Dayyán O'Brien , Bhavitvya Malik , Ona de Gibert , Pinzhen Chen , Barry Haddow , Jörg Tiedemann

The medical image is characterized by the inter-class indistinction, high variability, and noise, where the recognition of pixels is challenging. Unlike previous self-attention based methods that capture context information from one level,…

Computer Vision and Pattern Recognition · Computer Science 2019-11-26 Fei Ding , Gang Yang , Jinlu Liu , Jun Wu , Dayong Ding , Jie Xv , Gangwei Cheng , Xirong Li

High-resolution (HR) land-cover mapping is often constrained by the high cost of dense HR annotations. We revisit this problem from the perspective of map super-resolution, which enhances coarse low-resolution (LR) land-cover products into…

Computer Vision and Pattern Recognition · Computer Science 2026-04-17 Ruiqi Wang , Qi Yu , Jie Ma , Hanlin Wu

Document-level relation extraction aims at inferring structured human knowledge from textual documents. State-of-the-art methods for this task use pre-trained language models (LMs) via fine-tuning, yet fine-tuning is computationally…

Computation and Language · Computer Science 2024-10-03 Yilmazcan Ozyurt , Stefan Feuerriegel , Ce Zhang

Data is being produced in larger quantities than ever before in human history. It's only natural to expect a rise in demand for technology that aids humans in sifting through and analyzing this inexhaustible supply of information. This need…

Computation and Language · Computer Science 2020-02-13 Michael Kuehne , Marius Radu

Night time semantic segmentation is a crucial task in computer vision, focusing on accurately classifying and segmenting objects in low-light conditions. Unlike daytime techniques, which often perform worse in nighttime scenes, it is…

Computer Vision and Pattern Recognition · Computer Science 2024-07-09 Sarah Elmahdy , Rodaina Hebishy , Ali Hamdi

Real-world image super-resolution is a practical image restoration problem that aims to obtain high-quality images from in-the-wild input, has recently received considerable attention with regard to its tremendous application potentials.…

Computer Vision and Pattern Recognition · Computer Science 2022-06-07 Hao Li , Jinghui Qin , Zhijing Yang , Pengxu Wei , Jinshan Pan , Liang Lin , Yukai Shi

Editing High Dynamic Range (HDR) environment maps using an inverse differentiable rendering architecture is a complex inverse problem due to the sparsity of relevant pixels and the challenges in balancing light sources and background. The…

Computer Vision and Pattern Recognition · Computer Science 2024-10-25 Antonio D'Orazio , Davide Sforza , Fabio Pellacini , Iacopo Masi

Low Dose Computed Tomography (LDCT) is clinically desirable due to the reduced radiation to patients. However, the quality of LDCT images is often sub-optimal because of the inevitable strong quantum noise. Inspired by their unprecedent…

Image and Video Processing · Electrical Eng. & Systems 2021-02-02 Ti Bai , Dan Nguyen , Biling Wang , Steve Jiang

Single image dehazing is a prerequisite which affects the performance of many computer vision tasks and has attracted increasing attention in recent years. However, most existing dehazing methods emphasize more on haze removal but less on…

Computer Vision and Pattern Recognition · Computer Science 2021-09-23 Yan Li , De Cheng , Jiande Sun , Dingwen Zhang , Nannan Wang , Xinbo Gao

Image dehazing is one of the important and popular topics in computer vision and machine learning. A reliable real-time dehazing method with reliable performance is highly desired for many applications such as autonomous driving, security…

Computer Vision and Pattern Recognition · Computer Science 2022-12-16 Ruoteng Li , Xiaoyi Zhang , Shaodi You , Yu Li

Many existing 3D semantic segmentation methods, deep learning in computer vision notably, claimed to achieve desired results on urban point clouds. Thus, it is significant to assess these methods quantitatively in diversified real-world…

Computer Vision and Pattern Recognition · Computer Science 2023-12-12 Maosu Li , Yijie Wu , Anthony G. O. Yeh , Fan Xue