中文
相关论文

相关论文: Sky2Ground: A Benchmark for Site Modeling under Va…

200 篇论文

Current methods for 3D reconstruction and environmental mapping frequently face challenges in achieving high precision, highlighting the need for practical and effective solutions. In response to this issue, our study introduces FlyNeRF, a…

机器人学 · 计算机科学 2024-04-22 Maria Dronova , Vladislav Cheremnykh , Alexey Kotcov , Aleksey Fedoseev , Dzmitry Tsetserukou

This work has been accepted by IEEE TGRS for publication. The majority of optical observations acquired via spaceborne earth imagery are affected by clouds. While there is numerous prior work on reconstructing cloud-covered information,…

图像与视频处理 · 电气工程与系统科学 2021-07-07 Patrick Ebel , Andrea Meraner , Michael Schmitt , Xiaoxiang Zhu

Cloud phase profiles are critical for numerical weather prediction (NWP), as they directly affect radiative transfer and precipitation processes. In this study, we present a benchmark dataset and a baseline framework for transforming…

计算机视觉与模式识别 · 计算机科学 2025-09-22 Chi Yang , Fu Wang , Xiaofei Yang , Hao Huang , Weijia Cao , Xiaowen Chu

Regularly updated and accurate land cover maps are essential for monitoring 14 of the 17 Sustainable Development Goals. Multispectral satellite imagery provide high-quality and valuable information at global scale that can be used to…

计算机视觉与模式识别 · 计算机科学 2020-12-08 Hamed Alemohammad , Kevin Booth

Real-world robots localize objects from natural-language instructions while scenes around them keep changing. Yet most of the existing 3D visual grounding (3DVG) method still assumes a reconstructed and up-to-date point cloud, an assumption…

计算机视觉与模式识别 · 计算机科学 2025-10-17 Miao Hu , Zhiwei Huang , Tai Wang , Jiangmiao Pang , Dahua Lin , Nanning Zheng , Runsen Xu

3D recovery from multi-stereo and stereo images, as an important application of the image-based perspective geometry, serves many applications in computer vision, remote sensing and Geomatics. In this chapter, the authors utilize the…

计算机视觉与模式识别 · 计算机科学 2021-06-29 Rongjun Qin , Shuang Song , Xiao Ling , Mostafa Elhashash

Aerial-to-ground image synthesis is an emerging and challenging problem that aims to synthesize a ground image from an aerial image. Due to the highly different layout and object representation between the aerial and ground images, existing…

计算机视觉与模式识别 · 计算机科学 2023-08-15 Jinhyun Jang , Taeyong Song , Kwanghoon Sohn

Analyzing the planet at scale with satellite imagery and machine learning is a dream that has been constantly hindered by the cost of difficult-to-access highly-representative high-resolution imagery. To remediate this, we introduce here…

图像与视频处理 · 电气工程与系统科学 2025-06-03 Julien Cornebise , Ivan Oršolić , Freddie Kalaitzis

Buildings' segmentation is a fundamental task in the field of earth observation and aerial imagery analysis. Most existing deep learning-based methods in the literature can be applied to a fixed or narrow-range spatial resolution imagery.…

计算机视觉与模式识别 · 计算机科学 2023-10-04 Hasan Nasrallah , Mustafa Shukor , Ali J. Ghandour

The drone navigation requires the comprehensive understanding of both visual and geometric information in the 3D world. In this paper, we present a Visual-Geometric Fusion Network(VGF-Net), a deep network for the fusion analysis of…

计算机视觉与模式识别 · 计算机科学 2021-04-08 Yilin Liu , Ke Xie , Hui Huang

We introduce UprightNet, a learning-based approach for estimating 2DoF camera orientation from a single RGB image of an indoor scene. Unlike recent methods that leverage deep learning to perform black-box regression from image to…

计算机视觉与模式识别 · 计算机科学 2019-08-21 Wenqi Xian , Zhengqi Li , Matthew Fisher , Jonathan Eisenmann , Eli Shechtman , Noah Snavely

In this paper, we construct a large-scale benchmark dataset for Ground-to-Aerial Video-based person Re-Identification, named G2A-VReID, which comprises 185,907 images and 5,576 tracklets, featuring 2,788 distinct identities. To our…

计算机视觉与模式识别 · 计算机科学 2024-09-04 Shizhou Zhang , Wenlong Luo , De Cheng , Qingchun Yang , Lingyan Ran , Yinghui Xing , Yanning Zhang

We propose to use deep convolutional neural networks to address the problem of cross-view image geolocalization, in which the geolocation of a ground-level query image is estimated by matching to georeferenced aerial images. We use…

计算机视觉与模式识别 · 计算机科学 2015-10-14 Scott Workman , Richard Souvenir , Nathan Jacobs

We present Sat2Sound, a unified multimodal framework for geospatial soundscape understanding, designed to predict and map the distribution of sounds across the Earth's surface. Existing methods for this task rely on paired satellite images…

计算机视觉与模式识别 · 计算机科学 2026-04-14 Subash Khanal , Srikumar Sastry , Aayush Dhakal , Adeel Ahmad , Abby Stylianou , Nathan Jacobs

We introduce AG-VPReID, a new large-scale dataset for aerial-ground video-based person re-identification (ReID) that comprises 6,632 subjects, 32,321 tracklets and over 9.6 million frames captured by drones (altitudes ranging from 15-120m),…

计算机视觉与模式识别 · 计算机科学 2025-03-18 Huy Nguyen , Kien Nguyen , Akila Pemasiri , Feng Liu , Sridha Sridharan , Clinton Fookes

We consider the problem of cross-view geo-localization. The primary challenge of this task is to learn the robust feature against large viewpoint changes. Existing benchmarks can help, but are limited in the number of viewpoints. Image…

计算机视觉与模式识别 · 计算机科学 2020-08-18 Zhedong Zheng , Yunchao Wei , Yi Yang

Visual grounding (VG) aims to localize target objects in an image based on natural language descriptions. In this paper, we propose AerialVG, a new task focusing on visual grounding from aerial views. Compared to traditional VG, AerialVG…

计算机视觉与模式识别 · 计算机科学 2025-10-09 Junli Liu , Qizhi Chen , Zhigang Wang , Yiwen Tang , Yiting Zhang , Chi Yan , Dong Wang , Xuelong Li , Bin Zhao

Reconstruction of a continuous surface of two-dimensional manifold from its raw, discrete point cloud observation is a long-standing problem. The problem is technically ill-posed, and becomes more difficult considering that various sensing…

计算机视觉与模式识别 · 计算机科学 2022-05-06 Zhangjin Huang , Yuxin Wen , Zihao Wang , Jinjuan Ren , Kui Jia

Cross-view UAV geolocalization is fundamentally a challenging large-scale image retrieval task, aiming to determine the geographic coordinates of Unmanned Aerial Vehicle (UAV) queries by matching them against an extensive geo-tagged…

计算机视觉与模式识别 · 计算机科学 2026-05-18 Bowen Liu , Pengyue Jia , Wanyu Wang , Derong Xu , Jiawei Cheng , Jiancheng Dong , Xiao Han , Zimo Zhao , Chao Zhang , Bowen Yu , Fangyu Hong , Xiangyu Zhao

Visual grounding (VG) aims to establish fine-grained alignment between vision and language. Ideally, it can be a testbed for vision-and-language models to evaluate their understanding of the images and texts and their reasoning abilities…

计算机视觉与模式识别 · 计算机科学 2023-07-24 Zhihong Chen , Ruifei Zhang , Yibing Song , Xiang Wan , Guanbin Li