中文
相关论文

相关论文: M3TR: A Generalist Model for Real-World HD Map Com…

200 篇论文

Recent advances in scene reconstruction have pushed toward highly realistic modeling of autonomous driving (AD) environments using 3D Gaussian splatting. However, the resulting reconstructions remain closely tied to the original…

计算机视觉与模式识别 · 计算机科学 2025-12-15 Polina Karpikova , Daniil Selikhanovych , Kirill Struminsky , Ruslan Musaev , Maria Golitsyna , Dmitry Baranchuk

Lighting understanding plays an important role in virtual object composition, including mobile augmented reality (AR) applications. Prior work often targets recovering lighting from the physical environment to support photorealistic AR…

计算机视觉与模式识别 · 计算机科学 2023-01-18 Yiqin Zhao , Sean Fanello , Tian Guo

Recent advancements in neural visual geometry, including transformer-based models such as VGGT and Pi3, have achieved impressive accuracy on 3D reconstruction tasks. However, their reliance on full attention makes them fundamentally limited…

计算机视觉与模式识别 · 计算机科学 2026-03-04 Leo Kaixuan Cheng , Abdus Shaikh , Ruofan Liang , Zhijie Wu , Yushi Guan , Nandita Vijaykumar

In autonomous driving, High Definition (HD) maps provide a complete lane model that is not limited by sensor range and occlusions. However, the generation and upkeep of HD maps involves periodic data collection and human annotations,…

计算机视觉与模式识别 · 计算机科学 2024-09-20 Michael Mink , Thomas Monninger , Steffen Staab

Most Neural Radiance Fields (NeRFs) exhibit limited generalization capabilities, which restrict their applicability in representing multiple scenes using a single model. To address this problem, existing generalizable NeRF methods simply…

计算机视觉与模式识别 · 计算机科学 2024-03-20 Ganlin Yang , Guoqiang Wei , Zhizheng Zhang , Yan Lu , Dong Liu

Point cloud completion aims to recover complete 3D geometry from partial observations caused by limited viewpoints and occlusions. Existing learning-based works, including 3D Convolutional Neural Network (CNN)-based, point-based, and…

计算机视觉与模式识别 · 计算机科学 2026-01-28 Jiangyuan Liu , Yuhao Zhao , Hongxuan Ma , Zhe Liu , Jian Wang , Wei Zou

We propose a methodology for lidar super-resolution with ground vehicles driving on roadways, which relies completely on a driving simulator to enhance, via deep learning, the apparent resolution of a physical lidar. To increase the…

机器人学 · 计算机科学 2020-04-14 Tixiao Shan , Jinkun Wang , Fanfei Chen , Paul Szenher , Brendan Englot

Jointly processing information from multiple sensors is crucial to achieving accurate and robust perception for reliable autonomous driving systems. However, current 3D perception research follows a modality-specific paradigm, leading to…

计算机视觉与模式识别 · 计算机科学 2023-08-16 Haiyang Wang , Hao Tang , Shaoshuai Shi , Aoxue Li , Zhenguo Li , Bernt Schiele , Liwei Wang

Real-world image restoration aims to restore high-quality (HQ) images from degraded low-quality (LQ) inputs captured under uncontrolled conditions. Existing methods typically depend on ground-truth (GT) supervision, assuming that GT…

计算机视觉与模式识别 · 计算机科学 2026-04-07 Fengyang Xiao , Peng Hu , Lei Xu , XingE Guo , Guanyi Qin , Yuqi Shen , Chengyu Fang , Rihan Zhang , Chunming He , Sina Farsiu

Augmented reality (AR) has gained increasingly attention from both research and industry communities. By overlaying digital information and content onto the physical world, AR enables users to experience the world in a more informative and…

计算机视觉与模式识别 · 计算机科学 2021-03-30 Rui Huang , Chuan Fang , Kejie Qiu , Le Cui , Zilong Dong , Siyu Zhu , Ping Tan

Real scans always miss partial geometries of objects due to the self-occlusions, external-occlusions, and limited sensor resolutions. Point cloud completion aims to refer the complete shapes for incomplete 3D scans of objects. Current deep…

计算机视觉与模式识别 · 计算机科学 2022-03-22 Yiming Ren , Peishan Cong , Xinge Zhu , Yuexin Ma

High-quality 3D streaming from multiple cameras is crucial for immersive experiences in many AR/VR applications. The limited number of views - often due to real-time constraints - leads to missing information and incomplete surfaces in the…

计算机视觉与模式识别 · 计算机科学 2026-03-06 Leif Van Holland , Domenic Zingsheim , Mana Takhsha , Hannah Dröge , Patrick Stotko , Markus Plack , Reinhard Klein

We propose a novel approach for 3D shape completion by synthesizing multi-view depth maps. While previous work for shape completion relies on volumetric representations, meshes, or point clouds, we propose to use multi-view depth maps from…

计算机视觉与模式识别 · 计算机科学 2019-09-24 Tao Hu , Zhizhong Han , Abhinav Shrivastava , Matthias Zwicker

Lane detection is a vital task for vehicles to navigate and localize their position on the road. To ensure reliable driving, lane detection models must have robust generalization performance in various road environments. However, despite…

计算机视觉与模式识别 · 计算机科学 2024-06-04 Daeun Lee , Minhyeok Heo , Jiwon Kim

Multi-view transformers such as DUSt3R are revolutionizing 3D vision by solving 3D tasks in a feed-forward manner. However, contrary to previous optimization-based pipelines, the inner mechanisms of multi-view transformers are unclear.…

计算机视觉与模式识别 · 计算机科学 2025-10-30 Michal Stary , Julien Gaubil , Ayush Tewari , Vincent Sitzmann

High-definition maps (HD maps) play a crucial role in the development, safety validation, and operation of highly automated vehicles. Efficiently collecting up-to-date sensor data from road segments and obtaining accurate maps from these…

计算机视觉与模式识别 · 计算机科学 2024-10-02 Robert Krajewski , Huijo Kim

In a world where autonomous driving cars are becoming increasingly more common, creating an adequate infrastructure for this new technology is essential. This includes building and labeling high-definition (HD) maps accurately and…

计算机视觉与模式识别 · 计算机科学 2020-06-02 Mahdi Elhousni , Yecheng Lyu , Ziming Zhang , Xinming Huang

In recent years, tremendous efforts have been made on document image rectification, but existing advanced algorithms are limited to processing restricted document images, i.e., the input images must incorporate a complete document. Once the…

计算机视觉与模式识别 · 计算机科学 2023-12-19 Hao Feng , Shaokai Liu , Jiajun Deng , Wengang Zhou , Houqiang Li

With the release of open source datasets such as nuPlan and Argoverse, the research around learning-based planners has spread a lot in the last years. Existing systems have shown excellent capabilities in imitating the human driver…

机器人学 · 计算机科学 2025-04-22 Cristian Gariboldi , Matteo Corno , Beng Jin

In autonomous driving, robust place recognition is critical for global localization and loop closure detection. While inter-modality fusion of camera and LiDAR data in multimodal place recognition (MPR) has shown promise in overcoming the…

计算机视觉与模式识别 · 计算机科学 2026-02-24 Jingyi Xu , Zhangshuo Qi , Zhongmiao Yan , Xuyu Gao , Qianyun Jiao , Songpengcheng Xia , Xieyuanli Chen , Ling Pei