中文
相关论文

相关论文: ZeD-MAP: Bundle Adjustment Guided Zero-Shot Depth …

200 篇论文

Simultaneous Localization and Mapping (SLAM) is essential for precise surgical interventions and robotic tasks in minimally invasive procedures. While recent advancements in 3D Gaussian Splatting (3DGS) have improved SLAM with high-quality…

计算机视觉与模式识别 · 计算机科学 2025-02-03 Yiming Huang , Beilei Cui , Long Bai , Zhen Chen , Jinlin Wu , Zhen Li , Hongbin Liu , Hongliang Ren

Radar-camera fusion methods have emerged as a cost-effective approach for 3D object detection but still lag behind LiDAR-based methods in performance. Recent works have focused on employing temporal fusion and Knowledge Distillation (KD)…

计算机视觉与模式识别 · 计算机科学 2025-09-23 Geonho Bang , Minjae Seong , Jisong Kim , Geunju Baek , Daye Oh , Junhyung Kim , Junho Koh , Jun Won Choi

State-of-the-art techniques for monocular camera reconstruction predominantly rely on the Structure from Motion (SfM) pipeline. However, such methods often yield reconstruction outcomes that lack crucial scale information, and over time,…

机器人学 · 计算机科学 2023-10-10 Chunge Bai , Ruijie Fu , Xiang Gao

At present, the anchor-based or anchor-free models that use LiDAR point clouds for 3D object detection use the center assigner strategy to infer the 3D bounding boxes. However, in a real world scene, the LiDAR can only acquire a limited…

计算机视觉与模式识别 · 计算机科学 2022-03-08 Ruiqi Ma , Chi Chen , Bisheng Yang , Deren Li , Haiping Wang , Yangzi Cong , Zongtian Hu

3D visual grounding (3DVG) aims to localize objects in a 3D scene based on natural language queries. In this work, we explore zero-shot 3DVG from multi-view images alone, without requiring any geometric supervision or object priors. We…

计算机视觉与模式识别 · 计算机科学 2026-02-04 Nikita Drozdov , Andrey Lemeshko , Nikita Gavrilov , Anton Konushin , Danila Rukhovich , Maksim Kolodiazhnyi

This paper proposes a novel UAV-to-Vehicle (U2V) channel model for sixth-generation (6G) intelligent sensing-communication integration, based on three-dimensional (3D) scatterer prediction. To explore the mapping relationship between…

信号处理 · 电气工程与系统科学 2026-05-14 Shuo Wang , Zengrui Han , Lu Bai , Xiang Cheng

3D reconstruction from a single image is a long-standing problem in computer vision. Learning-based methods address its inherent scale ambiguity by leveraging increasingly large labeled and unlabeled datasets, to produce geometric priors…

计算机视觉与模式识别 · 计算机科学 2024-09-17 Vitor Guizilini , Pavel Tokmakov , Achal Dave , Rares Ambrus

Recent advances in deep-learning based methods for image matching have demonstrated their superiority over traditional algorithms, enabling correspondence estimation in challenging scenes with significant differences in viewing angles,…

计算机视觉与模式识别 · 计算机科学 2025-12-17 Rahul Deshmukh , Avinash Kak

Conventional SLAM systems using visual or LiDAR data often struggle in poor lighting and severe weather. Although 4D radar is suited for such environments, its sparse and noisy point clouds hinder accurate odometry estimation, while the…

机器人学 · 计算机科学 2025-12-11 Zhiheng Li , Weihua Wang , Qiang Shen , Yichen Zhao , Zheng Fang

3D Gaussian Splatting is a powerful visual representation, providing high-quality and efficient 3D scene reconstruction, but it is crucially dependent on accurate camera poses typically obtained from computationally intensive processes like…

机器人学 · 计算机科学 2026-04-15 Daniel Yang , Jungseok Hong , John J. Leonard , Yogesh Girdhar

Recent advances in depth sensing technologies allow fast electronic maneuvering of the laser beam, as opposed to fixed mechanical rotations. This will enable future sensors, in principle, to vary in real-time the sampling pattern. We…

计算机视觉与模式识别 · 计算机科学 2022-05-23 Ilya Tcenov , Guy Gilboa

Non-repetitive solid-state LiDAR scanning leads to an extremely sparse measurement regime for detecting airborne UAVs: a small quadrotor at 10-25 m typically produces only 1-2 returns per scan, which is far below the point densities assumed…

机器人学 · 计算机科学 2026-03-13 Nivand Khosravi , Rodrigo Ventura , Meysam Basiri

Flood prediction is critical for emergency planning and response to mitigate human and economic losses. Traditional physics-based hydrodynamic models generate high-resolution flood maps using numerical methods requiring fine-grid…

Unified image restoration is a significantly challenging task in low-level vision. Existing methods either make tailored designs for specific tasks, limiting their generalizability across various types of degradation, or rely on training…

计算机视觉与模式识别 · 计算机科学 2026-03-10 Huaqiu Li , Yong Wang , Tongwen Huang , Hailang Huang , Haoqian Wang , Xiangxiang Chu

Over the past decade, there has been a significant increase in the use of Unmanned Aerial Vehicles (UAVs) to support a wide variety of missions, such as remote surveillance, vehicle tracking, and object detection. For problems involving…

计算机视觉与模式识别 · 计算机科学 2022-12-06 Suleyman Melih Portakal , Ahmet Alp Kindiroglu , Mahiye Uluyagmur Ozturk

Automated Aerial Triangulation (AAT), aiming to restore image pose and reconstruct sparse points simultaneously, plays a pivotal role in earth observation. With its rich research heritage spanning several decades in photogrammetry, AAT has…

计算机视觉与模式识别 · 计算机科学 2024-04-09 Zequan Chen , Jianping Li , Qusheng Li , Bisheng Yang , Zhen Dong

Efficient structural perception is essential for mapping and autonomous navigation on resource-constrained robots. Existing 3D methods are computationally prohibitive, while traditional 2D geometric approaches lack robustness. This paper…

机器人学 · 计算机科学 2026-04-21 Guanliang Li , Pedro Espinosa-Angulo , David Perez-Saura , Santiago Tapia-Fernandez

This paper presents OpenREALM, a real-time mapping framework for Unmanned Aerial Vehicles (UAVs). A camera attached to the onboard computer of a moving UAV is utilized to acquire high resolution image mosaics of a targeted area of interest.…

计算机视觉与模式识别 · 计算机科学 2020-09-23 Alexander Kern , Markus Bobbe , Yogesh Khedar , Ulf Bestmann

Optical imaging through turbid or heterogeneous environments (collectively referred to as complex media) is fundamentally challenged by scattering, which scrambles structured spatial and phase information. To address this, we propose a…

Zero-shot 3D Visual Grounding (3DVG) is a critical capability for open-world embodied AI. However, existing methods are fundamentally bottlenecked by the poor quality of open-vocabulary 3D proposals, suffering from inaccurate categories and…

计算机视觉与模式识别 · 计算机科学 2026-04-30 Yufei Yin , Jie Zheng , Qianke Meng , Zhou Yu , Minghao Chen , Jiajun Ding , Min Tan , Yuling Xi , Zhiwen Chen , Chengfei Lv