English
Related papers

Related papers: Sky2Ground: A Benchmark for Site Modeling under Va…

200 papers

Visual grounding in 3D is the key for embodied agents to localize language-referred objects in open-world environments. However, existing benchmarks are limited to indoor focus, single-platform constraints, and small scale. We introduce…

Computer Vision and Pattern Recognition · Computer Science 2025-12-02 Rong Li , Yuhao Dong , Tianshuai Hu , Ao Liang , Youquan Liu , Dongyue Lu , Liang Pan , Lingdong Kong , Junwei Liang , Ziwei Liu

Ground robots play a crucial role in inspection, exploration, rescue, and other applications. In recent years, advancements in LiDAR technology have made sensors more accurate, lightweight, and cost-effective. Therefore, researchers…

Robotics · Computer Science 2025-03-18 Yanpeng Jia , Shiyi Wang , Shiliang Shao , Yue Wang , Fu Zhang , Ting Wang

We present a novel multi-altitude camera pose estimation system, addressing the challenges of robust and accurate localization across varied altitudes when only considering sparse image input. The system effectively handles diverse…

Computer Vision and Pattern Recognition · Computer Science 2025-08-14 Yaxuan Li , Yewei Huang , Bijay Gaudel , Hamidreza Jafarnejadsani , Brendan Englot

We present a novel method for synthesizing both temporally and geometrically consistent street-view panoramic video from a single satellite image and camera trajectory. Existing cross-view synthesis approaches focus on images, while video…

Computer Vision and Pattern Recognition · Computer Science 2021-05-11 Zuoyue Li , Zhenqiang Li , Zhaopeng Cui , Rongjun Qin , Marc Pollefeys , Martin R. Oswald

This paper presents the BigEarthNet that is a new large-scale multi-label Sentinel-2 benchmark archive. The BigEarthNet consists of 590,326 Sentinel-2 image patches, each of which is a section of i) 120x120 pixels for 10m bands; ii) 60x60…

Computer Vision and Pattern Recognition · Computer Science 2019-11-26 Gencer Sumbul , Marcela Charfuelan , Begüm Demir , Volker Markl

Images captured under low-light conditions are often plagued by several challenges, including diminished contrast, increased noise, loss of fine details, and unnatural color reproduction. These factors can significantly hinder the…

Computer Vision and Pattern Recognition · Computer Science 2023-05-16 Miao Zhang , Yiqing Shen , Shenghui Zhong

We present the DeepGlobe 2018 Satellite Image Understanding Challenge, which includes three public competitions for segmentation, detection, and classification tasks on satellite images. Similar to other challenges in computer vision domain…

Computer Vision and Pattern Recognition · Computer Science 2019-05-15 Ilke Demir , Krzysztof Koperski , David Lindenbaum , Guan Pang , Jing Huang , Saikat Basu , Forest Hughes , Devis Tuia , Ramesh Raskar

In this work, we construct a large-scale dataset for Ground-to-Aerial Person Search, named G2APS, which contains 31,770 images of 260,559 annotated bounding boxes for 2,644 identities appearing in both of the UAVs and ground surveillance…

Computer Vision and Pattern Recognition · Computer Science 2023-08-25 Shizhou Zhang , Qingchun Yang , De Cheng , Yinghui Xing , Guoqiang Liang , Peng Wang , Yanning Zhang

Deep learning provides a powerful new approach to many computer vision tasks. Height prediction from aerial images is one of those tasks that benefited greatly from the deployment of deep learning which replaced old multi-view geometry…

Computer Vision and Pattern Recognition · Computer Science 2021-11-15 Elhousni Mahdi , Zhang Ziming , Huang Xinming

3D visual grounding is an emerging research area dedicated to making connections between the 3D physical world and natural language, which is crucial for achieving embodied intelligence. In this paper, we propose DASANet, a Dual…

Computer Vision and Pattern Recognition · Computer Science 2024-06-14 Yue Xu , Kaizhi Yang , Jiebo Luo , Xuejin Chen

Super-Resolution (SR) has advanced rapidly in recent years, with diffusion-based models achieving unprecedented fidelity at the cost of introducing new types of visual artifacts. While existing Image Quality Assessment (IQA) methods provide…

Computer Vision and Pattern Recognition · Computer Science 2026-05-21 Artem Borisov , Evgeney Bogatyrev , Khaled Abud , Dmitriy Vatolin

Multi-view anomaly detection aims to identify surface defects on complex objects using observations captured from multiple viewpoints. However, existing unsupervised methods often suffer from feature inconsistency arising from viewpoint…

Computer Vision and Pattern Recognition · Computer Science 2026-04-08 Letian Bai , Chengyu Tao , Juan Du

Multimodal 3D grounding has garnered considerable interest in Vision-Language Models (VLMs) \cite{yin2025spatial} for advancing spatial reasoning in complex environments. However, these models suffer from a severe "2D semantic bias" that…

Computer Vision and Pattern Recognition · Computer Science 2025-10-21 Yutong Zhong

The multi-modal remote sensing foundation model (MM-RSFM) has significantly advanced various Earth observation tasks, such as urban planning, environmental monitoring, and natural disaster management. However, most existing approaches…

Computer Vision and Pattern Recognition · Computer Science 2025-07-21 Yingying Zhang , Lixiang Ru , Kang Wu , Lei Yu , Lei Liang , Yansheng Li , Jingdong Chen

Accurately describing and detecting 2D and 3D keypoints is crucial to establishing correspondences across images and point clouds. Despite a plethora of learning-based 2D or 3D local feature descriptors and detectors having been proposed,…

Computer Vision and Pattern Recognition · Computer Science 2021-07-30 Bing Wang , Changhao Chen , Zhaopeng Cui , Jie Qin , Chris Xiaoxuan Lu , Zhengdi Yu , Peijun Zhao , Zhen Dong , Fan Zhu , Niki Trigoni , Andrew Markham

Spatio-Temporal Video Grounding (STVG) aims to localize target objects in videos based on natural language descriptions. Despite recent advances in Multimodal Large Language Models, a significant gap remains between current models and…

Computer Vision and Pattern Recognition · Computer Science 2025-11-24 Hong Gao , Jingyu Wu , Xiangkai Xu , Kangni Xie , Yunchen Zhang , Bin Zhong , Xurui Gao , Min-Ling Zhang

Generating accurate 3D models is a challenging problem that traditionally requires explicit learning from 3D datasets using supervised learning. Although recent advances have shown promise in learning 3D models from 2D images, these methods…

Computer Vision and Pattern Recognition · Computer Science 2024-02-05 Qijia Shen , Guangrun Wang

Estimating building height from satellite imagery poses significant challenges, especially when monocular images are employed, resulting in a loss of essential 3D information during imaging. This loss of spatial depth further complicates…

Computer Vision and Pattern Recognition · Computer Science 2024-11-15 Mahd Qureshi , Shayaan Chaudhry , Sana Jabba , Murtaza Taj

Weather forecasting is one of the cornerstones of meteorological work. In this paper, we present a new benchmark dataset named Weather2K, which aims to make up for the deficiencies of existing weather forecasting datasets in terms of…

Machine Learning · Computer Science 2023-02-22 Xun Zhu , Yutong Xiong , Ming Wu , Gaozhen Nie , Bin Zhang , Ziheng Yang

Rain degrades the visual quality of multi-view images, which are essential for 3D scene reconstruction, resulting in inaccurate and incomplete reconstruction results. Existing datasets often overlook two critical characteristics of real…

Computer Vision and Pattern Recognition · Computer Science 2025-11-11 Qianfeng Yang , Xiang Chen , Pengpeng Li , Qiyuan Guan , Guiyue Jin , Jiyu Jin