中文
相关论文

相关论文: CrossVIT-augmented Geospatial-Intelligence Visuali…

200 篇论文

High-definition (HD) semantic mapping of complex intersections poses significant challenges for traditional vehicle-based approaches due to occlusions and limited perspectives. This paper introduces a novel camera-LiDAR fusion framework…

机器人学 · 计算机科学 2025-07-15 Zhongzhang Chen , Miao Fan , Shengtong Xu , Mengmeng Yang , Kun Jiang , Xiangzeng Liu , Haoyi Xiong

Understanding urban socioeconomic conditions through visual data is a challenging yet essential task for sustainable urban development and policy planning. In this work, we introduce \textit{CityLens}, a comprehensive benchmark designed to…

人工智能 · 计算机科学 2026-03-03 Tianhui Liu , Hetian Pang , Xin Zhang , Tianjian Ouyang , Zhiyuan Zhang , Jie Feng , Yong Li , Pan Hui

Object navigation (ObjectNav) in real-world environments is a complex problem that requires simultaneously addressing multiple challenges, including complex spatial structure, long-horizon planning and semantic understanding. Recent…

机器人学 · 计算机科学 2026-03-10 Haokun Zhu , Zongtai Li , Zihan Liu , Kevin Guo , Zhengzhi Lin , Yuxin Cai , Guofei Chen , Chen Lv , Wenshan Wang , Jean Oh , Ji Zhang

Pervasive localization is essential for continuous tracking applications, yet existing solutions face challenges in balancing power consumption and accuracy. GPS, while precise, is impractical for continuous tracking of micro-assets due to…

系统与控制 · 电气工程与系统科学 2025-05-12 Aritrik Ghosh , Nakul Garg , Nirupam Roy

Recent self-supervised learning (SSL) methods have demonstrated impressive results in learning visual representations from unlabeled remote sensing images. However, most remote sensing images predominantly consist of scenographic scenes…

计算机视觉与模式识别 · 计算机科学 2024-11-12 Kaixuan Lu , Ruiqian Zhang , Xiao Huang , Yuxing Xie , Xiaogang Ning , Hanchao Zhang , Mengke Yuan , Pan Zhang , Tao Wang , Tongkui Liao

Ever more robust, accurate and detailed mapping using visual sensing has proven to be an enabling factor for mobile robots across a wide variety of applications. For the next level of robot intelligence and intuitive user interaction, maps…

计算机视觉与模式识别 · 计算机科学 2016-09-29 John McCormac , Ankur Handa , Andrew Davison , Stefan Leutenegger

Cross-modal retrieval aims to measure the content similarity between different types of data. The idea has been previously applied to visual, text, and speech data. In this paper, we present a novel cross-modal retrieval method specifically…

计算机视觉与模式识别 · 计算机科学 2020-05-05 Numan Khurshid , Talha Hanif , Mohbat Tharani , Murtaza Taj

Visual localization is the problem of estimating the position and orientation from which a given image (or a sequence of images) is taken in a known scene. It is an important part of a wide range of computer vision and robotics…

计算机视觉与模式识别 · 计算机科学 2021-09-13 Ara Jafarzadeh , Manuel Lopez Antequera , Pau Gargallo , Yubin Kuang , Carl Toft , Fredrik Kahl , Torsten Sattler

How to economically cluster large-scale multi-view images is a long-standing problem in computer vision. To tackle this challenge, we introduce a novel approach named Highly-economized Scalable Image Clustering (HSIC) that radically…

计算机视觉与模式识别 · 计算机科学 2018-09-18 Zheng Zhang , Li Liu , Jie Qin , Fan Zhu , Fumin Shen , Yong Xu , Ling Shao , Heng Tao Shen

In this paper, we present TANDEM a real-time monocular tracking and dense mapping framework. For pose estimation, TANDEM performs photometric bundle adjustment based on a sliding window of keyframes. To increase the robustness, we propose a…

计算机视觉与模式识别 · 计算机科学 2021-11-16 Lukas Koestler , Nan Yang , Niclas Zeller , Daniel Cremers

Earth observation offers new insight into anthropogenic changes to nature, and how these changes are effecting (and are effected by) the built environment and the real economy. With the global availability of medium-resolution (10-30m)…

计算机视觉与模式识别 · 计算机科学 2021-02-15 Lucas Kruitwagen

Semantic segmentation aims to robustly predict coherent class labels for entire regions of an image. It is a scene understanding task that powers real-world applications (e.g., autonomous navigation). One important application, the use of…

计算机视觉与模式识别 · 计算机科学 2023-02-16 Yuxiang Zhang , Sachin Mehta , Anat Caspi

We present a novel multi-altitude camera pose estimation system, addressing the challenges of robust and accurate localization across varied altitudes when only considering sparse image input. The system effectively handles diverse…

计算机视觉与模式识别 · 计算机科学 2025-08-14 Yaxuan Li , Yewei Huang , Bijay Gaudel , Hamidreza Jafarnejadsani , Brendan Englot

Dense prediction tasks hold significant importance of computer vision, aiming to learn pixel-wise annotated labels for input images. Despite advances in this field, existing methods primarily focus on idealized conditions, exhibiting…

计算机视觉与模式识别 · 计算机科学 2025-10-01 Changliang Xia , Chengyou Jia , Zhuohang Dang , Minnan Luo , Zhihui Li , Xiaojun Chang

Recently, Referring Remote Sensing Image Segmentation (RRSIS) has aroused wide attention. To handle drastic scale variation of remote targets, existing methods only use the full image as input and nest the saliency-preferring techniques of…

计算机视觉与模式识别 · 计算机科学 2025-08-05 Jiaxing Yang , Lihe Zhang , Huchuan Lu

Visual analytics is essential for studying large time series due to its ability to reveal trends, anomalies, and insights. DeepVATS is a tool that merges Deep Learning (Deep) with Visual Analytics (VA) for the analysis of large time series…

机器学习 · 计算机科学 2025-04-02 Inmaculada Santamaria-Valenzuela , Victor Rodriguez-Fernandez , David Camacho

At modern construction sites, utilizing GNSS (Global Navigation Satellite System) to measure the real-time location and orientation (i.e. pose) of construction machines and navigate them is very common. However, GNSS is not always…

机器人学 · 计算机科学 2021-01-19 Runqiu Bao , Ren Komatsu , Renato Miyagusuku , Masaki Chino , Atsushi Yamashita , Hajime Asama

We present a fast, spatio-temporal scene understanding framework based on Visual Geometry Grounded Transformer (VGGT). The proposed pipeline is designed to enable efficient, close to real-time performance, supporting applications including…

计算机视觉与模式识别 · 计算机科学 2025-12-01 Gergely Dinya , Péter Halász , András Lőrincz , Kristóf Karacs , Anna Gelencsér-Horváth

OdoViz is a reactive web-based tool for 3D visualization and processing of autonomous vehicle datasets designed to support common tasks in visual place recognition research. The system includes functionality for loading, inspecting,…

计算机视觉与模式识别 · 计算机科学 2021-07-19 Saravanabalagi Ramachandran , John McDonald

Explainability in time series forecasting is essential for improving model transparency and supporting informed decision-making. In this work, we present CrossScaleNet, an innovative architecture that combines a patch-based cross-attention…

计算机视觉与模式识别 · 计算机科学 2025-09-30 Ibrahim Delibasoglu , Fredrik Heintz
‹ 上一页 1 8 9 10 下一页 ›