中文
相关论文

相关论文: FM-Loc: Using Foundation Models for Improved Visio…

200 篇论文

Visual localization is the problem of estimating a camera within a scene and a key component in computer vision applications such as self-driving cars and Mixed Reality. State-of-the-art approaches for accurate visual localization use…

计算机视觉与模式识别 · 计算机科学 2021-03-30 Qunjie Zhou , Torsten Sattler , Marc Pollefeys , Laura Leal-Taixe

Place recognition is a challenging problem in mobile robotics, especially in unstructured environments or under viewpoint and illumination changes. Most LiDAR-based methods rely on geometrical features to overcome such challenges, as…

机器人学 · 计算机科学 2018-12-03 Jiadong Guo , Paulo V. K. Borges , Chanoh Park , Abel Gawel

Recent learning-based approaches have achieved impressive results in the field of single-shot camera localization. However, how best to fuse multiple modalities (e.g., image and depth) and to deal with degraded or missing input are less…

计算机视觉与模式识别 · 计算机科学 2025-01-22 Kaichen Zhou , Changhao Chen , Bing Wang , Muhamad Risqi U. Saputra , Niki Trigoni , Andrew Markham

Siamese tracking paradigm has achieved great success, providing effective appearance discrimination and size estimation by the classification and regression. While such a paradigm typically optimizes the classification and regression…

计算机视觉与模式识别 · 计算机科学 2022-05-02 Jiahao Nie , Han Wu , Zhiwei He , Yuxiang Yang , Mingyu Gao , Zhekang Dong

Accurate robot localization is essential for effective operation. Monte Carlo Localization (MCL) is commonly used with known maps but is computationally expensive due to landmark matching for each particle. Humanoid robots face additional…

机器人学 · 计算机科学 2025-05-19 Ruochen Hou , Mingzhang Zhu , Hyunwoo Nam , Gabriel I. Fernandez , Dennis W. Hong

Multi-label recognition with partial labels (MLR-PL), in which only some labels are known while others are unknown for each image, is a practical task in computer vision, since collecting large-scale and complete multi-label datasets is…

计算机视觉与模式识别 · 计算机科学 2024-12-17 Haoxian Ruan , Zhihua Xu , Zhijing Yang , Yongyi Lu , Jinghui Qin , Tianshui Chen

This paper proposes Spatially-Weighted CLIP (SW-CLIP), a novel framework for street-view geo-localization that explicitly incorporates spatial autocorrelation into vision-language contrastive learning. Unlike conventional CLIP-based methods…

计算机视觉与模式识别 · 计算机科学 2026-04-07 Ting Han , Fengjiao Li , Chunsong Chen , Haoling Huang , Yiping Chen , Meiliu Wu

We seek to predict the 6 degree-of-freedom (6DoF) pose of a query photograph with respect to a large indoor 3D map. The contributions of this work are three-fold. First, we develop a new large-scale visual localization method targeted for…

计算机视觉与模式识别 · 计算机科学 2018-04-10 Hajime Taira , Masatoshi Okutomi , Torsten Sattler , Mircea Cimpoi , Marc Pollefeys , Josef Sivic , Tomas Pajdla , Akihiko Torii

Capturing human mobility is essential for modeling how people interact with and move through physical spaces, reflecting social behavior, access to resources, and dynamic spatial patterns. To support scalable and transferable analysis…

人工智能 · 计算机科学 2025-06-18 Mohammad Hashemi , Andreas Zufle

Intelligent robots require object-level scene understanding to reason about possible tasks and interactions with the environment. Moreover, many perception tasks such as scene reconstruction, image retrieval, or place recognition can…

计算机视觉与模式识别 · 计算机科学 2023-05-05 Cathrin Elich , Iro Armeni , Martin R. Oswald , Marc Pollefeys , Joerg Stueckler

We present GSplatLoc, a camera localization method that leverages the differentiable rendering capabilities of 3D Gaussian splatting for ultra-precise pose estimation. By formulating pose estimation as a gradient-based optimization problem…

计算机视觉与模式识别 · 计算机科学 2025-05-20 Atticus J. Zeller , Haijuan Wu

Robust robot localization is an important prerequisite for navigation, but it becomes challenging when the map and robot measurements are obtained from different sensors. Prior methods are often tailored to specific environments, relying on…

机器人学 · 计算机科学 2026-04-03 Evgenii Kruzhkov , Raphael Memmesheimer , Sven Behnke

Feature matching is crucial in visual localization, where 2D-3D correspondence plays a major role in determining the accuracy of camera pose. A sufficient number of well-distributed 2D-3D correspondences is essential for accurate pose…

计算机视觉与模式识别 · 计算机科学 2023-03-07 Hailin Yu , Youji Feng , Weicai Ye , Mingxuan Jiang , Hujun Bao , Guofeng Zhang

We introduce a high-fidelity neural implicit dense visual Simultaneous Localization and Mapping (SLAM) system, termed DF-SLAM. In our work, we employ dictionary factors for scene representation, encoding the geometry and appearance…

计算机视觉与模式识别 · 计算机科学 2024-06-27 Weifeng Wei , Jie Wang , Shuqi Deng , Jie Liu

Simultaneous Localization And Mapping (SLAM) is a fundamental problem in mobile robotics. While point-based SLAM methods provide accurate camera localization, the generated maps lack semantic information. On the other hand, state of the art…

机器人学 · 计算机科学 2018-11-05 Mehdi Hosseinzadeh , Yasir Latif , Trung Pham , Niko Suenderhauf , Ian Reid

Event-based cameras offer much potential to the fields of robotics and computer vision, in part due to their large dynamic range and extremely high "frame rates". These attributes make them, at least in theory, particularly suitable for…

机器人学 · 计算机科学 2015-05-19 Michael Milford , Hanme Kim , Michael Mangan , Stefan Leutenegger , Tom Stone , Barbara Webb , Andrew Davison

Versatile and adaptive semantic understanding would enable autonomous systems to comprehend and interact with their surroundings. Existing fixed-class models limit the adaptability of indoor mobile and assistive autonomous systems. In this…

机器人学 · 计算机科学 2024-03-06 Christina Kassab , Matias Mattamala , Lintong Zhang , Maurice Fallon

While large-scale image-text pretrained models such as CLIP have been used for multiple video-level tasks on trimmed videos, their use for temporal localization in untrimmed videos is still a relatively unexplored task. We design a new…

计算机视觉与模式识别 · 计算机科学 2023-08-23 Shen Yan , Xuehan Xiong , Arsha Nagrani , Anurag Arnab , Zhonghao Wang , Weina Ge , David Ross , Cordelia Schmid

Accurate localisation is critical for mobile robots in structured outdoor environments, yet LiDAR-based methods often fail in vineyards due to repetitive row geometry and perceptual aliasing. We propose a semantic particle filter that…

机器人学 · 计算机科学 2025-09-24 Rajitha de Silva , Jonathan Cox , James R. Heselden , Marija Popovic , Cesar Cadena , Riccardo Polvara

Vision-Language Models (VLMs), such as CLIP, have demonstrated remarkable zero-shot generalizability across diverse downstream tasks. However, recent studies have revealed that VLMs, including CLIP, are highly vulnerable to adversarial…

密码学与安全 · 计算机科学 2025-10-27 Jia Deng , Jin Li , Zhenhua Zhao , Shaowei Wang
‹ 上一页 1 8 9 10 下一页 ›