English
Related papers

Related papers: A model-based approach for transforming InSAR-deri…

200 papers

Predicting temporal progress from visual trajectories is important for intelligent robots that can learn, adapt, and improve. However, learning such progress estimator, or temporal value function, across different tasks and domains requires…

Radar presents a promising alternative to lidar and vision in autonomous vehicle applications, able to detect objects at long range under a variety of weather conditions. However, distinguishing between occupied and free space from raw…

Robotics · Computer Science 2019-05-13 Rob Weston , Sarah Cen , Paul Newman , Ingmar Posner

Synthetic Aperture Radar (SAR) is a crucial remote sensing technology, enabling all-weather, day-and-night observation with strong surface penetration for precise and continuous environmental monitoring and analysis. However, SAR image…

Computer Vision and Pattern Recognition · Computer Science 2025-04-07 Yimin Wei , Aoran Xiao , Yexian Ren , Yuting Zhu , Hongruixuan Chen , Junshi Xia , Naoto Yokoya

Global point cloud registration is an essential module for localization, of which the main difficulty exists in estimating the rotation globally without initial value. With the aid of gravity alignment, the degree of freedom in point cloud…

Robotics · Computer Science 2022-03-03 Xiaqing Ding , Xuecheng Xu , Sha Lu , Yanmei Jiao , Mengwen Tan , Rong Xiong , Huanjun Deng , Mingyang Li , Yue Wang

Vision-Language Models (VLMs) excel at 2D tasks such as grounding and captioning, yet remain limited in 3D understanding. A key limitation is their text-only supervision paradigm, which under-constrains fine-grained visual perception and…

Computer Vision and Pattern Recognition · Computer Science 2026-05-21 Hanxun Yu , Xuan Qu , Yuxin Wang , Jianke Zhu , Lei Ke

Visual Place Recognition (VPR) has evolved from handcrafted descriptors to deep learning approaches, yet significant challenges remain. Current approaches, including Vision Foundation Models (VFMs) and Multimodal Large Language Models…

Machine Learning · Computer Science 2025-09-03 Jintao Cheng , Weibin Li , Jiehao Luo , Xiaoyu Tang , Zhijian He , Jin Wu , Yao Zou , Wei Zhang

Open-set perception in complex traffic environments poses a critical challenge for autonomous driving systems, particularly in identifying previously unseen object categories, which is vital for ensuring safety. Visual Language Models…

Computer Vision and Pattern Recognition · Computer Science 2025-08-13 Fuhao Chang , Shuxin Li , Yabei Li , Lei He

Trained on internet-scale video data, generative world models are increasingly recognized as powerful world simulators that can generate consistent and plausible dynamics over structure, motion, and physics. This raises a natural question:…

Computer Vision and Pattern Recognition · Computer Science 2025-10-02 Kevin Zhang , Kuangzhi Ge , Xiaowei Chi , Renrui Zhang , Shaojun Shi , Zhen Dong , Sirui Han , Shanghang Zhang

We present a novel Simultaneous Localization and Mapping (SLAM) method that employs Gaussian Process (GP) based landmark (object) representations. Instead of conventional grid maps or point cloud registration, we model the environment on a…

Robotics · Computer Science 2025-08-25 Ali Emre Balcı , Erhan Ege Keyvan , Emre Özkan

Near-field extremely large multiple input multiple output (XL-MIMO) breaks the assumptions that make classical super-resolution effective: the receiver acquires only a limited set of compressed pilot observations, while each propagation…

Signal Processing · Electrical Eng. & Systems 2026-04-14 Sajad Daei , Gabor Fodor , Mikael Skoglund

Vision-based sensors have shown significant performance, accuracy, and efficiency gain in Simultaneous Localization and Mapping (SLAM) systems in recent years. In this regard, Visual Simultaneous Localization and Mapping (VSLAM) methods…

Computer Vision and Pattern Recognition · Computer Science 2022-12-02 Ali Tourani , Hriday Bavle , Jose Luis Sanchez-Lopez , Holger Voos

Recent advancements in autonomous driving, augmented reality, robotics, and embodied intelligence have necessitated 3D perception algorithms. However, current 3D perception methods, especially specialized small models, exhibit poor…

Computer Vision and Pattern Recognition · Computer Science 2025-02-14 Fan Yang , Sicheng Zhao , Yanhao Zhang , Hui Chen , Haonan Lu , Jungong Han , Guiguang Ding

Using geometric landmarks like lines and planes can increase navigation accuracy and decrease map storage requirements compared to commonly-used LiDAR point cloud maps. However, landmark-based registration for applications like loop closure…

Robotics · Computer Science 2022-12-27 Parker C. Lusk , Devarth Parikh , Jonathan P. How

With advancements in data availability and computing resources, Multimodal Large Language Models (MLLMs) have showcased capabilities across various fields. However, the quadratic complexity of the vision encoder in MLLMs constrains the…

Computer Vision and Pattern Recognition · Computer Science 2024-07-24 Yiwei Ma , Zhibin Wang , Xiaoshuai Sun , Weihuang Lin , Qiang Zhou , Jiayi Ji , Rongrong Ji

Accurate localization is a core component of a robot's navigation system. To this end, global navigation satellite systems (GNSS) can provide absolute measurements outdoors and, therefore, eliminate long-term drift. However, fusing GNSS…

Robotics · Computer Science 2024-10-10 Jonas Beuchert , Marco Camurri , Maurice Fallon

While recent large vision-language models (VLMs) have improved generalization in vision-language navigation (VLN), existing methods typically rely on end-to-end pipelines that map vision-language inputs directly to short-horizon discrete…

3D terrain reconstruction with remote sensing imagery achieves cost-effective and large-scale earth observation and is crucial for safeguarding natural disasters, monitoring ecological changes, and preserving the environment.Recently,…

Computer Vision and Pattern Recognition · Computer Science 2025-01-03 Song Zhang , Zhiwei Wei , Wenjia Xu , Lili Zhang , Yang Wang , Jinming Zhang , Junyi Liu

Global climate models (GCMs), typically run at ~100-km resolution, capture large-scale environmental conditions but cannot resolve convection and cloud processes at kilometer scales. Convection-permitting models offer higher-resolution…

Atmospheric and Oceanic Physics · Physics 2026-05-12 Hungjui Yu , Lander Ver Hoef , Kristen L. Rasmussen , Imme Ebert-Uphoff

The rapid advancement of Large Multimodal Models (LMMs) for 2D images and videos has motivated extending these models to understand 3D scenes, aiming for human-like visual-spatial intelligence. Nevertheless, achieving deep spatial…

Interferometric Synthetic Aperture Radar (InSAR) technology uses satellite radar to detect surface deformation patterns and monitor earthquake impacts on buildings. While vital for emergency response planning, extracting multi-class…

Computer Vision and Pattern Recognition · Computer Science 2025-02-27 Xuechun Li , Susu Xu
‹ Prev 1 8 9 10 Next ›