English
Related papers

Related papers: Geo-ID: Test-Time Geometric Consensus for Cross-Vi…

200 papers

While InfoNCE underlies modern contrastive learning, its geometric mechanisms remain under-characterized beyond the canonical alignment--uniformity decomposition. We develop a measure-theoretic framework in which representation measures…

Machine Learning · Computer Science 2026-05-26 Yichao Cai , Zhen Zhang , Yuhang Liu , Javen Qinfeng Shi

Intrinsic Image Decomposition is an open problem of generating the constituents of an image. Generating reflectance and shading from a single image is a challenging task specifically when there is no ground truth. There is a lack of…

Computer Vision and Pattern Recognition · Computer Science 2021-12-21 Harshana Weligampola , Gihan Jayatilaka , Suren Sritharan , Parakrama Ekanayake , Roshan Ragel , Vijitha Herath , Roshan Godaliyadda

Exploiting internal spatial geometric constraints of sparse LiDARs is beneficial to depth completion, however, has been not explored well. This paper proposes an efficient method to learn geometry-aware embedding, which encodes the local…

Computer Vision and Pattern Recognition · Computer Science 2022-06-02 Wenchao Du , Hu Chen , Hongyu Yang , Yi Zhang

Recent advances in Gaussian Splatting-based inverse rendering extend Gaussian primitives with shading parameters and physically grounded light transport, enabling high-quality material recovery from dense multi-view captures. However, these…

Computer Vision and Pattern Recognition · Computer Science 2025-12-11 Patrick Noras , Jun Myeong Choi , Didier Stricker , Pieter Peers , Roni Sengupta

We present a method for finding cross-modal space-time correspondences. Given two images from different visual modalities, such as an RGB image and a depth map, our model identifies which pairs of pixels correspond to the same physical…

Computer Vision and Pattern Recognition · Computer Science 2025-06-04 Ayush Shrivastava , Andrew Owens

Reconstructing coherent 3D geometry and appearance from unposed multi-view images is a fundamental yet challenging problem in computer vision. Most existing visual geometry foundation models predict explicit geometry by regressing…

Computer Vision and Pattern Recognition · Computer Science 2026-05-22 Yuqi Wu , Tianyu Hu , Wenzhao Zheng , Yuanhui Huang , Haowen Sun , Jie Zhou , Jiwen Lu

3D Gaussians have recently emerged as an effective scene representation for real-time splatting and accurate novel-view synthesis, motivating several works to adapt multi-view structure prediction networks to regress per-pixel 3D Gaussians…

Computer Vision and Pattern Recognition · Computer Science 2025-12-22 Mehdi Hosseinzadeh , Shin-Fang Chng , Yi Xu , Simon Lucey , Ian Reid , Ravi Garg

Recent advances in 3D Gaussian Splatting have enabled impressive photorealistic novel view synthesis. However, to transition from a pure rendering engine to a reliable spatial map for autonomous agents and safety-critical applications,…

Computer Vision and Pattern Recognition · Computer Science 2026-03-25 Chamuditha Jayanga Galappaththige , Thomas Gottwald , Peter Stehr , Edgar Heinert , Niko Suenderhauf , Dimity Miller , Matthias Rottmann

Using accurate depth priors in 3D Gaussian Splatting helps mitigate artifacts caused by sparse training data and textureless surfaces. However, acquiring accurate depth maps requires specialized acquisition systems. Foundation monocular…

Computer Vision and Pattern Recognition · Computer Science 2026-04-08 Wenhui Xiao , Ethan Goan , Rodrigo Santa Cruz , David Ahmedt-Aristizabal , Olivier Salvado , Clinton Fookes , Leo Lebrat

Although 3D Gaussian Splatting has been widely studied because of its realistic and efficient novel-view synthesis, it is still challenging to extract a high-quality surface from the point-based representation. Previous works improve the…

Computer Vision and Pattern Recognition · Computer Science 2024-10-31 Hanlin Chen , Fangyin Wei , Chen Li , Tianxin Huang , Yunsong Wang , Gim Hee Lee

3D Gaussian Splatting (3DGS) creates a radiance field consisting of 3D Gaussians to represent a scene. With sparse training views, 3DGS easily suffers from overfitting, negatively impacting rendering. This paper introduces a new…

Computer Vision and Pattern Recognition · Computer Science 2024-07-12 Jiawei Zhang , Jiahe Li , Xiaohan Yu , Lei Huang , Lin Gu , Jin Zheng , Xiao Bai

Neural multivariate regression underpins a wide range of domains, including control, robotics, and finance, yet the geometry of its learned representations remains poorly characterized. While neural collapse has been shown to benefit…

Machine Learning · Computer Science 2026-05-11 George Andriopoulos , Zixuan Dong , Bimarsha Adhikari , Keith Ross

Advancing towards artificial superintelligence requires rich and intelligent perceptual capabilities. A critical frontier in this pursuit is overcoming the limited spatial understanding of Multimodal Large Language Models (MLLMs), where…

Computer Vision and Pattern Recognition · Computer Science 2026-03-12 Ruiheng Liu , Haihong Hao , Mingfei Han , Xin Gu , Kecheng Zhang , Changlin Li , Xiaojun Chang

Co-saliency detection (Co-SOD) aims to segment the common salient foreground in a group of relevant images. In this paper, inspired by human behavior, we propose a gradient-induced co-saliency detection (GICD) method. We first abstract a…

Computer Vision and Pattern Recognition · Computer Science 2020-12-15 Zhao Zhang , Wenda Jin , Jun Xu , Ming-Ming Cheng

This paper addresses the solution of inverse problems in imaging given an additional reference image. We combine a modification of the discrete geodesic path model for image metamorphosis with a variational model,actually the $L^2$-$TV$…

Numerical Analysis · Mathematics 2019-05-22 Sebastian Neumayer , Johannes Persch , Gabriele Steidl

Recently, 3D Gaussian Splatting (3D-GS) has achieved significant success in real-time, high-quality 3D scene rendering. However, it faces several challenges, including Gaussian redundancy, limited ability to capture view-dependent effects,…

Computer Vision and Pattern Recognition · Computer Science 2025-05-16 Junxi Jin , Xiulai Li , Haiping Huang , Lianjun Liu , Yujie Sun , Logan Liu

Scene view synthesis, which generates novel views from limited perspectives, is increasingly vital for applications like virtual reality, augmented reality, and robotics. Unlike object-based tasks, such as generating 360{\deg} views of a…

Computer Vision and Pattern Recognition · Computer Science 2025-04-29 Xiaofeng Jin , Yan Fang , Matteo Frosi , Jianfei Ge , Jiangjian Xiao , Matteo Matteucci

Multimodal synthetic data generation is crucial in domains such as autonomous driving, robotics, augmented/virtual reality, and retail. We propose a novel approach, GenMM, for jointly editing RGB videos and LiDAR scans by inserting…

Computer Vision and Pattern Recognition · Computer Science 2024-06-18 Bharat Singh , Viveka Kulharia , Luyu Yang , Avinash Ravichandran , Ambrish Tyagi , Ashish Shrivastava

Monocular depth inference is a fundamental problem for scene perception of robots. Specific robots may be equipped with a camera plus an optional depth sensor of any type and located in various scenes of different scales, whereas recent…

Computer Vision and Pattern Recognition · Computer Science 2023-10-25 Haotian Wang , Meng Yang , Nanning Zheng

We present a joint 3D pose and focal length estimation approach for object categories in the wild. In contrast to previous methods that predict 3D poses independently of the focal length or assume a constant focal length, we explicitly…

Computer Vision and Pattern Recognition · Computer Science 2019-08-09 Alexander Grabner , Peter M. Roth , Vincent Lepetit