中文
相关论文

相关论文: UVDoc: Neural Grid-based Document Unwarping

200 篇论文

Variational methods are widely applied to ill-posed inverse problems for they have the ability to embed prior knowledge about the solution. However, the level of performance of these methods significantly depends on a set of parameters,…

Unsigned Distance Fields (UDFs) provide a flexible representation for 3D shapes with arbitrary topology, including open and closed surfaces, orientable and non-orientable geometries, and non-manifold structures. While recent neural…

计算机视觉与模式识别 · 计算机科学 2025-10-15 Jiayi Kong , Chen Zong , Junkai Deng , Xuhui Chen , Fei Hou , Shiqing Xin , Junhui Hou , Chen Qian , Ying He

Capturing the shape and spatially-varying appearance (SVBRDF) of an object from images is a challenging task that has applications in both computer vision and graphics. Traditional optimization-based approaches often need a large number of…

计算机视觉与模式识别 · 计算机科学 2021-05-20 Mark Boss , Varun Jampani , Kihwan Kim , Hendrik P. A. Lensch , Jan Kautz

With the development of deep learning, medical image processing has been widely used to assist clinical research. This paper focuses on the denoising problem of low-dose computed tomography using deep learning. Although low-dose computed…

图像与视频处理 · 电气工程与系统科学 2026-05-19 Zhilin Guan , Wei Zhang

Recent image restoration methods can be broadly categorized into two classes: (1) regression methods that recover the rough structure of the original image without synthesizing high-frequency details and (2) generative methods that…

计算机视觉与模式识别 · 计算机科学 2024-01-02 Hwayoon Lee , Kyoungkook Kang , Hyeongmin Lee , Seung-Hwan Baek , Sunghyun Cho

In this paper, we introduce a novel task termed unified anomaly detection and classification, which aims to simultaneously detect anomalous regions in images and identify their specific categories. Existing methods typically treat anomaly…

计算机视觉与模式识别 · 计算机科学 2026-02-04 Ximiao Zhang , Min Xu , Zheng Zhang , Junlin Hu , Xiuzhuang Zhou

The protection of intellectual property has become critical due to the rapid growth of three-dimensional content in digital media. Unlike traditional images or videos, 3D point clouds present unique challenges for copyright enforcement, as…

计算机视觉与模式识别 · 计算机科学 2026-04-07 Khandoker Ashik Uz Zaman , Mohammad Zahangir Alam , Mohammed N. M. Ali , Mahdi H. Miraz

Deep unfolding networks have gained increasing attention in the field of compressed sensing (CS) owing to their theoretical interpretability and superior reconstruction performance. However, most existing deep unfolding methods often face…

图像与视频处理 · 电气工程与系统科学 2025-04-17 Kai Han , Jin Wang , Yunhui Shi , Hanqin Cai , Nam Ling , Baocai Yin

We present recursive cascaded networks, a general architecture that enables learning deep cascades, for deformable image registration. The proposed architecture is simple in design and can be built on any base network. The moving image is…

计算机视觉与模式识别 · 计算机科学 2020-03-26 Shengyu Zhao , Yue Dong , Eric I-Chao Chang , Yan Xu

Limited data and low dose constraints are common problems in a variety of tomographic reconstruction paradigms which lead to noisy and incomplete data. Over the past few years sinogram denoising has become an essential pre-processing step…

计算机视觉与模式识别 · 计算机科学 2016-03-15 Faisal Mahmood , Nauman Shahid , Pierre Vandergheynst , Ulf Skoglund

We present an unsupervised data-driven approach for non-rigid shape matching. Shape matching identifies correspondences between two shapes and is a fundamental step in many computer vision and graphics applications. Our approach is designed…

计算机视觉与模式识别 · 计算机科学 2023-11-28 Aymen Merrouche , Joao Regateiro , Stefanie Wuhrer , Edmond Boyer

In contrast to traditional image restoration methods, all-in-one image restoration techniques are gaining increased attention for their ability to restore images affected by diverse and unknown corruption types and levels. However,…

计算机视觉与模式识别 · 计算机科学 2024-01-25 Yimin Xu , Nanxi Gao , Zhongyun Shan , Fei Chao , Rongrong Ji

Deep neural networks (DNNs), while increasingly deployed in many applications, struggle with robustness against anomalous and out-of-distribution (OOD) data. Current OOD benchmarks often oversimplify, focusing on single-object tasks and not…

计算机视觉与模式识别 · 计算机科学 2024-09-30 Debargha Ganguly , Debayan Gupta , Vipin Chaudhary

Rigid image alignment is a fundamental task in computer vision, while the traditional algorithms are either too sensitive to noise or time-consuming. Recent unsupervised image alignment methods developed based on spatial transformer…

计算机视觉与模式识别 · 计算机科学 2022-05-25 Yu-Xuan Chen , Dagan Feng , Hong-Bin Shen

3D visual grounding allows an embodied agent to understand visual information in real-world 3D environments based on human instructions, which is crucial for embodied intelligence. Existing 3D visual grounding methods typically rely on…

计算机视觉与模式识别 · 计算机科学 2025-09-05 Fan Li , Zanyi Wang , Zeyi Huang , Guang Dai , Jingdong Wang , Mengmeng Wang

We present an approach to matching images of objects in fine-grained datasets without using part annotations, with an application to the challenging problem of weakly supervised single-view reconstruction. This is in contrast to prior works…

计算机视觉与模式识别 · 计算机科学 2016-06-21 Angjoo Kanazawa , David W. Jacobs , Manmohan Chandraker

Recent advances in data-centric deep generative models have led to significant progress in solving inverse imaging problems. However, these models (e.g., diffusion models (DMs)) typically require large amounts of fully sampled (clean)…

计算机视觉与模式识别 · 计算机科学 2025-05-20 Shijun Liang , Ismail R. Alkhouri , Siddhant Gautam , Qing Qu , Saiprasad Ravishankar

Degraded document image binarization is one of the most challenging tasks in the domain of document image analysis. In this paper, we present a novel approach towards document image binarization by introducing three-player min-max…

计算机视觉与模式识别 · 计算机科学 2020-10-28 Amandeep Kumar , Shuvozit Ghose , Pinaki Nath Chowdhury , Partha Pratim Roy , Umapada Pal

Learning non-rigid registration in an end-to-end manner is challenging due to the inherent high degrees of freedom and the lack of labeled training data. In this paper, we resolve these two challenges simultaneously. First, we propose to…

计算机视觉与模式识别 · 计算机科学 2021-04-14 Wanquan Feng , Juyong Zhang , Hongrui Cai , Haofei Xu , Junhui Hou , Hujun Bao

Since real-time contents can be captured and downloaded very easily, copyright infringement has become a serious problem. In order to reduce the loss caused by copyright infringement, copyright owners insert a watermark in the content to…

多媒体 · 计算机科学 2018-05-17 Wook-Hyung Kim , Jong-Uk Hou , Seung-Min Mun , Heung-Kyu Lee