中文
相关论文

相关论文: SQS: Enhancing Sparse Perception Models via Query-…

200 篇论文

We propose DrivingForward, a feed-forward Gaussian Splatting model that reconstructs driving scenes from flexible surround-view input. Driving scene images from vehicle-mounted cameras are typically sparse, with limited overlap, and the…

计算机视觉与模式识别 · 计算机科学 2024-12-24 Qijian Tian , Xin Tan , Yuan Xie , Lizhuang Ma

Traditional SLAM algorithms excel at camera tracking, but typically produce incomplete and low-resolution maps that are not tightly integrated with semantics prediction. Recent work integrates Gaussian Splatting (GS) into SLAM to enable…

计算机视觉与模式识别 · 计算机科学 2025-12-04 Mingqi Jiang , Chanho Kim , Chen Ziwen , Li Fuxin

Cooperative perception is critical for autonomous driving, overcoming the inherent limitations of a single vehicle, such as occlusions and constrained fields-of-view. However, current approaches sharing dense Bird's-Eye-View (BEV) features…

计算机视觉与模式识别 · 计算机科学 2026-04-13 Jiahao Wang , Zhongwei Jiang , Wenchao Sun , Jiaru Zhong , Haibao Yu , Yuner Zhang , Chenyang Lu , Chuang Zhang , Lei He , Shaobing Xu , Jianqiang Wang

Sensor simulation is pivotal for scalable validation of autonomous driving systems, yet existing Neural Radiance Fields (NeRF) based methods face applicability and efficiency challenges in industrial workflows. This paper introduces a…

3D Gaussian splatting (3DGS) is an innovative rendering technique that surpasses the neural radiance field (NeRF) in both rendering speed and visual quality by leveraging an explicit 3D scene representation. Existing 3DGS approaches require…

计算机视觉与模式识别 · 计算机科学 2025-06-13 Lintao Xiang , Hongpei Zheng , Yating Huang , Qijun Yang , Hujun Yin

Vision-based autonomous driving has gained much attention due to its low costs and excellent performance. Compared with dense BEV (Bird's Eye View) or sparse query models, Gaussian-centric method is a comprehensive yet sparse representation…

计算机视觉与模式识别 · 计算机科学 2026-04-02 Yiyao Zhu , Ying Xue , Haiming Zhang , Guangfeng Jiang , Wending Zhou , Xu Yan , Jiantao Gao , Yingjie Cai , Bingbing Liu , Zhen Li , Shaojie Shen

Self-supervised learning of point cloud aims to leverage unlabeled 3D data to learn meaningful representations without reliance on manual annotations. However, current approaches face challenges such as limited data diversity and inadequate…

计算机视觉与模式识别 · 计算机科学 2024-09-10 Keyi Liu , Yeqi Luo , Weidong Yang , Jingyi Xu , Zhijun Li , Wen-Ming Chen , Ben Fei

Recently, 3D Gaussian Splatting (3DGS) provides a new framework for novel view synthesis, and has spiked a new wave of research in neural rendering and related applications. As 3DGS is becoming a foundational component of many models, any…

计算机视觉与模式识别 · 计算机科学 2025-04-14 Jialin Zhu , Jiangbei Yue , Feixiang He , He Wang

We propose Self-Augmented Residual 3D Gaussian Splatting (SA-ResGS), a novel framework to stabilize uncertainty quantification and enhancing uncertainty-aware supervision in next-best-view (NBV) selection for active scene reconstruction.…

计算机视觉与模式识别 · 计算机科学 2026-01-21 Kim Jun-Seong , Tae-Hyun Oh , Eduardo Pérez-Pellitero , Youngkyoon Jang

The well-established modular autonomous driving system is decoupled into different standalone tasks, e.g. perception, prediction and planning, suffering from information loss and error accumulation across modules. In contrast, end-to-end…

计算机视觉与模式识别 · 计算机科学 2024-06-03 Wenchao Sun , Xuewu Lin , Yining Shi , Chuang Zhang , Haoran Wu , Sifa Zheng

Realistic scene reconstruction and view synthesis are essential for advancing autonomous driving systems by simulating safety-critical scenarios. 3D Gaussian Splatting excels in real-time rendering and static scene reconstructions but…

计算机视觉与模式识别 · 计算机科学 2024-07-08 Mustafa Khan , Hamidreza Fazlali , Dhruv Sharma , Tongtong Cao , Dongfeng Bai , Yuan Ren , Bingbing Liu

3D Gaussian Splatting (3DGS) has demonstrated remarkable real-time performance in novel view synthesis, yet its effectiveness relies heavily on dense multi-view inputs with precisely known camera poses, which are rarely available in…

计算机视觉与模式识别 · 计算机科学 2025-08-22 Zongqi He , Hanmin Li , Kin-Chung Chan , Yushen Zuo , Hao Xie , Zhe Xiao , Jun Xiao , Kin-Man Lam

We propose GGS, a Generalizable Gaussian Splatting method for Autonomous Driving which can achieve realistic rendering under large viewpoint changes. Previous generalizable 3D gaussian splatting methods are limited to rendering novel views…

计算机视觉与模式识别 · 计算机科学 2024-09-05 Huasong Han , Kaixuan Zhou , Xiaoxiao Long , Yusen Wang , Chunxia Xiao

We present SparseGen, a novel framework for efficient image-to-3D generation, which exhibits low input-view bias while being significantly faster. Unlike traditional approaches that rely on dense volumetric grids, triplanes, or…

计算机视觉与模式识别 · 计算机科学 2026-04-16 Zhiyuan Xu , Jiuming Liu , Yuxin Chen , Masayoshi Tomizuka , Chenfeng Xu , Chensheng Peng

Accurately predicting 3D occupancy grids from visual inputs is critical for autonomous driving, but current discriminative methods struggle with noisy data, incomplete observations, and the complex structures inherent in 3D scenes. In this…

计算机视觉与模式识别 · 计算机科学 2025-07-04 Yunshen Wang , Yicheng Liu , Tianyuan Yuan , Yingshi Liang , Xiuyu Yang , Honggang Zhang , Hang Zhao

3D Language Gaussian Splatting (3DLGS) augments 3D Gaussian Splatting with language-aligned visual features for open-vocabulary 3D scene understanding. A core challenge is efficiently associating high-dimensional vision-language embeddings…

计算机视觉与模式识别 · 计算机科学 2026-05-14 Lovre Antonio Budimir , Yushi Guan , Steve Ryhner , Sven Lončarić , Nandita Vijaykumar

Recently, several studies have combined Gaussian Splatting to obtain scene representations with language embeddings for open-vocabulary 3D scene understanding. While these methods perform well, they essentially require very dense multi-view…

计算机视觉与模式识别 · 计算机科学 2024-12-05 Jun Hu , Zhang Chen , Zhong Li , Yi Xu , Juyong Zhang

Monocular object pose estimation, as a pivotal task in computer vision and robotics, heavily depends on accurate 2D-3D correspondences, which often demand costly CAD models that may not be readily available. Object 3D reconstruction methods…

计算机视觉与模式识别 · 计算机科学 2024-09-05 Luqing Luo , Shichu Sun , Jiangang Yang , Linfang Zheng , Jinwei Du , Jian Liu

Vision-based perception for autonomous driving requires an explicit modeling of a 3D space, where 2D latent representations are mapped and subsequent 3D operators are applied. However, operating on dense latent spaces introduces a cubic…

计算机视觉与模式识别 · 计算机科学 2024-04-16 Pin Tang , Zhongdao Wang , Guoqing Wang , Jilai Zheng , Xiangxuan Ren , Bailan Feng , Chao Ma

We present a Gaussian Splatting method for surface reconstruction using sparse input views. Previous methods relying on dense views struggle with extremely sparse Structure-from-Motion points for initialization. While learning-based…

计算机视觉与模式识别 · 计算机科学 2025-04-30 Jiang Wu , Rui Li , Yu Zhu , Rong Guo , Jinqiu Sun , Yanning Zhang