English
Related papers

Related papers: PMatch: Paired Masked Image Modeling for Dense Geo…

200 papers

Three-dimensional Morphable Models (3DMMs) are powerful statistical tools for representing the 3D surfaces of an object class. In this context, we identify an interesting question that has previously not received research attention: is it…

Computer Vision and Pattern Recognition · Computer Science 2019-03-12 Stylianos Ploumpis , Haoyang Wang , Nick Pears , William A. P. Smith , Stefanos Zafeiriou

Recent image inpainting methods have made great progress but often struggle to generate plausible image structures when dealing with large holes in complex images. This is partially due to the lack of effective network structures that can…

Computer Vision and Pattern Recognition · Computer Science 2022-07-22 Haitian Zheng , Zhe Lin , Jingwan Lu , Scott Cohen , Eli Shechtman , Connelly Barnes , Jianming Zhang , Ning Xu , Sohrab Amirghodsi , Jiebo Luo

Establishing point-to-point correspondences across multiple 3D shapes is a fundamental problem in computer vision and graphics. In this paper, we introduce DcMatch, a novel unsupervised learning framework for non-rigid multi-shape matching.…

Computer Vision and Pattern Recognition · Computer Science 2025-11-13 Tianwei Ye , Yong Ma , Xiaoguang Mei

We propose an efficient approach to train large diffusion models with masked transformers. While masked transformers have been extensively explored for representation learning, their application to generative learning is less explored in…

Computer Vision and Pattern Recognition · Computer Science 2024-03-06 Hongkai Zheng , Weili Nie , Arash Vahdat , Anima Anandkumar

Many Multi-View-Stereo algorithms extract a 3D mesh model of a scene, after fusing depth maps into a volumetric representation of the space. Due to the limited scalability of such representations, the estimated model does not capture fine…

Computer Vision and Pattern Recognition · Computer Science 2019-05-23 Andrea Romanoni , Matteo Matteucci

Establishing consistent correspondences across images is essential for 3D vision tasks such as structure-from-motion (SfM), yet most existing matchers operate in a pairwise manner, often producing fragmented and geometrically inconsistent…

Computer Vision and Pattern Recognition · Computer Science 2026-03-31 Jongmin Lee , Seungyeop Kang , Sungjoo Yoo

How do the neural networks distinguish two images? It is of critical importance to understand the matching mechanism of deep models for developing reliable intelligent systems for many risky visual applications such as surveillance and…

Computer Vision and Pattern Recognition · Computer Science 2021-08-13 Wenliang Zhao , Yongming Rao , Ziyi Wang , Jiwen Lu , Jie Zhou

Supervised and unsupervised homography estimation methods depend on image pairs tailored to specific modalities to achieve high accuracy. However, their performance deteriorates substantially when applied to unseen modalities. To address…

Computer Vision and Pattern Recognition · Computer Science 2026-03-05 Jinkun You , Jiaxin Cheng , Jie Zhang , Yicong Zhou

Recent unified image generation models have achieved remarkable success by employing MLLMs for semantic understanding and diffusion backbones for image generation. However, these models remain fundamentally limited in spatially-aware tasks…

Computer Vision and Pattern Recognition · Computer Science 2026-04-30 Haiyi Qiu , Kaihang Pan , Jiacheng Li , Juncheng Li , Siliang Tang , Yueting Zhuang

Finding correspondences between shapes is a fundamental problem in computer vision and graphics, which is relevant for many applications, including 3D reconstruction, object tracking, and style transfer. The vast majority of correspondence…

Computer Vision and Pattern Recognition · Computer Science 2024-04-04 Maolin Gao , Zorah Lähner , Johan Thunberg , Daniel Cremers , Florian Bernard

One-shot face re-enactment is a challenging task due to the identity mismatch between source and driving faces. Specifically, the suboptimally disentangled identity information of driving subjects would inevitably interfere with the…

Computer Vision and Pattern Recognition · Computer Science 2022-11-24 Yunfan Liu , Qi Li , Zhenan Sun , Tieniu Tan

In computer vision, finding correct point correspondence among images plays an important role in many applications, such as image stitching, image retrieval, visual localization, etc. Most of the research works focus on the matching of…

Computer Vision and Pattern Recognition · Computer Science 2023-05-30 Yueh-Cheng Huang , Chen-Tao Hsu , Jen-Hui Chuang

Establishing correspondences across images is a fundamental challenge in computer vision, underpinning tasks like Structure-from-Motion, image editing, and point tracking. Traditional methods are often specialized for specific…

Computer Vision and Pattern Recognition · Computer Science 2025-01-28 Fei Xue , Sven Elflein , Laura Leal-Taixé , Qunjie Zhou

The rapid advancement of Artificial Intelligence (AI) in biomedical imaging and radiotherapy is hindered by the limited availability of large imaging data repositories. With recent research and improvements in denoising diffusion…

Image and Video Processing · Electrical Eng. & Systems 2024-03-27 Rowan Bradbury , Katherine A. Vallis , Bartlomiej W. Papiez

Although synthetic data can alleviate acquisition challenges in image dehazing tasks, it also introduces the problem of domain bias when dealing with small-scale data. This paper proposes a novel dual-branch collaborative unpaired dehazing…

Computer Vision and Pattern Recognition · Computer Science 2024-07-16 Shuaibin Fan , Minglong Xue , Aoxiang Ning , Senming Zhong

While the keypoint-based maps created by sparse monocular simultaneous localisation and mapping (SLAM) systems are useful for camera tracking, dense 3D reconstructions may be desired for many robotic tasks. Solutions involving depth cameras…

Computer Vision and Pattern Recognition · Computer Science 2022-07-26 Tristan Laidlow , Jan Czarnowski , Stefan Leutenegger

Masking strategies commonly employed in natural language processing are still underexplored in vision tasks such as concept learning, where conventional methods typically rely on full images. However, using masked images diversifies…

Computer Vision and Pattern Recognition · Computer Science 2025-11-14 Yuwei Sun , Lu Mi , Ippei Fujisawa , Ruiqiao Mei , Jimin Chen , Siyu Zhu , Ryota Kanai

Although large-scale visual foundation models (VFMs) achieve remarkable performance in semantic understanding, they still underperform in instance-aware dense prediction tasks. They exhibit different biases in representation: for instance,…

Computer Vision and Pattern Recognition · Computer Science 2026-05-19 Yachan Guo , JoseLuis Gomez Zurita , Danna Xue , Yi Xiao , AntonioManuel Lopez Pena

Image matting is generally modeled as a space transform from the color space to the alpha space. By estimating the alpha factor of the model, the foreground of an image can be extracted. However, there is some dimensional information…

Computer Vision and Pattern Recognition · Computer Science 2019-04-17 Xuelong Li , Kang Liu , Yongsheng Dong , Dacheng Tao

This paper presents iMatcher, a fully differentiable framework for feature matching in point cloud registration. The proposed method leverages learned features to predict a geometrically consistent confidence matrix, incorporating both…

Computer Vision and Pattern Recognition · Computer Science 2025-09-12 Karim Slimani , Catherine Achard , Brahim Tamadazte