中文
相关论文

相关论文: BevSplat: Resolving Height Ambiguity via Feature-B…

200 篇论文

While most recent autonomous driving system focuses on developing perception methods on ego-vehicle sensors, people tend to overlook an alternative approach to leverage intelligent roadside cameras to extend the perception ability beyond…

计算机视觉与模式识别 · 计算机科学 2023-04-12 Lei Yang , Kaicheng Yu , Tao Tang , Jun Li , Kun Yuan , Li Wang , Xinyu Zhang , Peng Chen

Accurate BEV semantic segmentation from fisheye imagery remains challenging due to extreme non-linear distortion, occlusion, and depth ambiguity inherent to wide-angle projections. We present a distortion-aware BEV segmentation framework…

计算机视觉与模式识别 · 计算机科学 2025-11-24 Shubham Sonarghare , Prasad Deshpande , Ciaran Hogan , Deepika-Rani Kaliappan-Mahalingam , Ganesh Sistu

Recently, generalizable human Gaussian splatting from sparse-view inputs has been actively studied for the photorealistic human rendering. Most existing methods rely on explicit geometric constraints or predefined structural representations…

计算机视觉与模式识别 · 计算机科学 2026-04-29 Jingi Kim , Wonjun Kim

We present ViewSplat, a view-adaptive 3D Gaussian splatting network for novel view synthesis from unposed images. While recent feed-forward 3D Gaussian splatting has significantly accelerated 3D scene reconstruction by bypassing per-scene…

计算机视觉与模式识别 · 计算机科学 2026-03-27 Moonyeon Jeong , Seunggi Min , Suhyeon Lee , Hongje Seong

Cross-view geo-localization for Unmanned Aerial Vehicles (UAVs) operating in GNSS-denied environments remains challenging due to the severe geometric discrepancy between oblique UAV imagery and orthogonal satellite maps. Most existing…

计算机视觉与模式识别 · 计算机科学 2026-04-03 Haoyuan Li , Wen Yang , Fang Xu , Hong Tan , Haijian Zhang , Shengyang Li , Gui-Song Xia

This paper proposes a fine-grained self-localization method for outdoor robotics that utilizes a flexible number of onboard cameras and readily accessible satellite images. The proposed method addresses limitations in existing cross-view…

计算机视觉与模式识别 · 计算机科学 2023-08-17 Shan Wang , Yanhao Zhang , Akhil Perincherry , Ankit Vora , Hongdong Li

Ground to aerial matching is a crucial and challenging task in outdoor robotics, particularly when GPS is absent or unreliable. Structures like buildings or large dense forests create interference, requiring GNSS replacements for global…

机器人学 · 计算机科学 2024-10-10 Christopher Klammer , Michael Kaess

This work addresses visual cross-view metric localization for outdoor robotics. Given a ground-level color image and a satellite patch that contains the local surroundings, the task is to identify the location of the ground camera within…

计算机视觉与模式识别 · 计算机科学 2022-08-19 Zimin Xia , Olaf Booij , Marco Manfredi , Julian F. P. Kooij

We present Cross-View Splatter, a feed-forward method that predicts pixel-aligned Gaussian splats for outdoor scenes captured at ground level AND by satellite. Faithful reconstructions require good camera coverage, but ground imagery is…

Gaussian Splatting has achieved remarkable progress in multi-view surface reconstruction, yet it exhibits notable degradation when only few views are available. Although recent efforts alleviate this issue by enhancing multi-view…

计算机视觉与模式识别 · 计算机科学 2026-05-13 Jimin Tang , Wenyuan Zhang , Junsheng Zhou , Zian Huang , Kanle Shi , Shenkun Xu , Yu-Shen Liu , Zhizhong Han

We propose an accurate and interpretable fine-grained cross-view localization method that estimates the 3 Degrees of Freedom (DoF) pose of a ground-level image by matching its local features with a reference aerial image. Unlike prior…

计算机视觉与模式识别 · 计算机科学 2026-02-27 Zimin Xia , Chenghao Xu , Alexandre Alahi

Depth estimation is a cornerstone of perception in autonomous driving and robotic systems. The considerable cost and relatively sparse data acquisition of LiDAR systems have led to the exploration of cost-effective alternatives, notably,…

计算机视觉与模式识别 · 计算机科学 2023-06-21 Yucheng Mao , Ruowen Zhao , Tianbao Zhang , Hang Zhao

This paper investigates the advantages of using Bird's Eye View (BEV) representation in 360-degree visual place recognition (VPR). We propose a novel network architecture that utilizes the BEV representation in feature extraction, feature…

计算机视觉与模式识别 · 计算机科学 2023-05-24 Xuecheng Xu , Yanmei Jiao , Sha Lu , Xiaqing Ding , Rong Xiong , Yue Wang

We introduce NoPoSplat, a feed-forward model capable of reconstructing 3D scenes parameterized by 3D Gaussians from \textit{unposed} sparse multi-view images. Our model, trained exclusively with photometric loss, achieves real-time 3D…

计算机视觉与模式识别 · 计算机科学 2024-11-01 Botao Ye , Sifei Liu , Haofei Xu , Xueting Li , Marc Pollefeys , Ming-Hsuan Yang , Songyou Peng

We present VicaSplat, a novel framework for joint 3D Gaussians reconstruction and camera pose estimation from a sequence of unposed video frames, which is a critical yet underexplored task in real-world 3D applications. The core of our…

计算机视觉与模式识别 · 计算机科学 2025-03-14 Zhiqi Li , Chengrui Dong , Yiming Chen , Zhangchi Huang , Peidong Liu

Recent advances in 3D Gaussian Splatting have enabled impressive photorealistic novel view synthesis. However, to transition from a pure rendering engine to a reliable spatial map for autonomous agents and safety-critical applications,…

计算机视觉与模式识别 · 计算机科学 2026-03-25 Chamuditha Jayanga Galappaththige , Thomas Gottwald , Peter Stehr , Edgar Heinert , Niko Suenderhauf , Dimity Miller , Matthias Rottmann

Gaussian Splatting demonstrates impressive results in multi-view reconstruction based on Gaussian explicit representations. However, the current Gaussian primitives only have a single view-dependent color and an opacity to represent the…

计算机视觉与模式识别 · 计算机科学 2026-05-05 Rui Xu , Wenyue Chen , Jiepeng Wang , Yuan Liu , Peng Wang , Cheng Lin , Shiqing Xin , Xin Li , Wenping Wang , Taku Komura

We introduce SPFSplatV2, an efficient feed-forward framework for 3D Gaussian splatting from sparse multi-view images, requiring no ground-truth poses during training and inference. It employs a shared feature extraction backbone, enabling…

计算机视觉与模式识别 · 计算机科学 2025-09-23 Ranran Huang , Krystian Mikolajczyk

Accurate 3D lane detection from monocular images presents significant challenges due to depth ambiguity and imperfect ground modeling. Previous attempts to model the ground have often used a planar ground assumption with limited degrees of…

计算机视觉与模式识别 · 计算机科学 2025-01-27 Chaesong Park , Eunbin Seo , Jongwoo Lim

A major breakthrough in 3D reconstruction is the feedforward paradigm to generate pixel-wise 3D points or Gaussian primitives from sparse, unposed images. To further incorporate semantics while avoiding the significant memory and storage…

计算机视觉与模式识别 · 计算机科学 2025-10-13 Yu Sheng , Jiajun Deng , Xinran Zhang , Yu Zhang , Bei Hua , Yanyong Zhang , Jianmin Ji