English
Related papers

Related papers: CityRefer: Geography-aware 3D Visual Grounding Dat…

200 papers

Many existing 3D semantic segmentation methods, deep learning in computer vision notably, claimed to achieve desired results on urban point clouds. Thus, it is significant to assess these methods quantitatively in diversified real-world…

Computer Vision and Pattern Recognition · Computer Science 2023-12-12 Maosu Li , Yijie Wu , Anthony G. O. Yeh , Fan Xue

Deep learning approaches have shown promising results in remote sensing high spatial resolution (HSR) land-cover mapping. However, urban and rural scenes can show completely different geographical landscapes, and the inadequate…

Computer Vision and Pattern Recognition · Computer Science 2022-06-01 Junjue Wang , Zhuo Zheng , Ailong Ma , Xiaoyan Lu , Yanfei Zhong

Autonomous vehicles generate massive volumes of point cloud data, yet only a subset is relevant for specific tasks such as collision detection, traffic analysis, or congestion monitoring. Effectively querying this data is essential to…

Computer Vision and Pattern Recognition · Computer Science 2025-08-04 Xiaoyu Zhang , Zhifeng Bao , Hai Dong , Ziwei Wang , Jiajun Liu

Synthetic datasets, recognized for their cost effectiveness, play a pivotal role in advancing computer vision tasks and techniques. However, when it comes to remote sensing image processing, the creation of synthetic datasets becomes…

Computer Vision and Pattern Recognition · Computer Science 2023-09-06 Jian Song , Hongruixuan Chen , Naoto Yokoya

We tackle the problem of localizing 3D point cloud submaps using complex and diverse natural language descriptions, and present Text2Loc++, a novel neural network designed for effective cross-modal alignment between language and point…

Computer Vision and Pattern Recognition · Computer Science 2025-11-20 Yan Xia , Letian Shi , Yilin Di , Joao F. Henriques , Daniel Cremers

Accurate 3D reconstruction of objects with reflective, transparent, or low-texture surfaces still remains notoriously challenging. Such materials often violate key assumptions in multi-view reconstruction pipelines, such as photometric…

Computer Vision and Pattern Recognition · Computer Science 2026-05-12 Zhicheng Liang , Haoyi Yu , Boyan Li , Dayou Zhang , Zijian Cao , Tianyi Gong , Junhua Liu , Shuguang Cui , Fangxin Wang

3D visual grounding consists of identifying the instance in a 3D scene which is referred by an accompanying language description. While several architectures have been proposed within the commonly employed grounding-by-selection framework,…

Computer Vision and Pattern Recognition · Computer Science 2024-11-07 Sombit Dey , Ozan Unal , Christos Sakaridis , Luc Van Gool

We address the problem of generating a 3D-consistent, navigable environment that is spatially grounded: a simulation of a real location. Existing video generative models can produce a plausible sequence that is consistent with a text (T2V)…

Computer Vision and Pattern Recognition · Computer Science 2026-04-22 Gene Chou , Charles Herrmann , Kyle Genova , Boyang Deng , Songyou Peng , Bharath Hariharan , Jason Y. Zhang , Noah Snavely , Philipp Henzler

Localizing objects in 3D scenes according to the semantics of a given natural language is a fundamental yet important task in the field of multimedia understanding, which benefits various real-world applications such as robotics and…

Computer Vision and Pattern Recognition · Computer Science 2023-09-06 Wencan Huang , Daizong Liu , Wei Hu

Data collection for autonomous driving is rapidly accelerating, but manual annotation, especially for 3D labels, remains a major bottleneck due to its high cost and labor intensity. Autolabeling has emerged as a scalable alternative,…

Computer Vision and Pattern Recognition · Computer Science 2025-07-29 Levente Tempfli , Esteban Rivera , Markus Lienkamp

We introduce the UT Campus Object Dataset (CODa), a mobile robot egocentric perception dataset collected on the University of Texas Austin Campus. Our dataset contains 8.5 hours of multimodal sensor data: synchronized 3D point clouds and…

We introduce RaidaR, a rich annotated image dataset of rainy street scenes, to support autonomous driving research. The new dataset contains the largest number of rainy images (58,542) to date, 5,000 of which provide semantic segmentations…

Computer Vision and Pattern Recognition · Computer Science 2021-10-27 Jiongchao Jin , Arezou Fatemi , Wallace Lira , Fenggen Yu , Biao Leng , Rui Ma , Ali Mahdavi-Amiri , Hao Zhang

Geo-localizing static objects from street images is challenging but also very important for road asset mapping and autonomous driving. In this paper we present a two-stage framework that detects and geolocalizes traffic signs from low frame…

Computer Vision and Pattern Recognition · Computer Science 2021-07-14 Daniel Wilson , Thayer Alshaabi , Colin Van Oort , Xiaohan Zhang , Jonathan Nelson , Safwan Wshah

Large-scale semantic mapping is crucial for outdoor autonomous agents to fulfill high-level tasks such as planning and navigation. This paper proposes a novel method for large-scale 3D semantic reconstruction through implicit…

Computer Vision and Pattern Recognition · Computer Science 2024-03-21 Jianyuan Zhang , Zhiliu Yang , Meng Zhang

Building recognition and 3D reconstruction of human made structures in urban scenarios has become an interesting and actual topic in the image processing domain. For this research topic the Computer Vision and Augmented Reality areas…

Computer Vision and Pattern Recognition · Computer Science 2021-10-28 Orhei Ciprian , Vert Silviu , Mocofan Muguras , Vasiu Radu

With the acceleration of the urban expansion, urban change detection (UCD), as a significant and effective approach, can provide the change information with respect to geospatial objects for dynamical urban analysis. However, existing…

Computer Vision and Pattern Recognition · Computer Science 2020-12-29 Shiqi Tian , Ailong Ma , Zhuo Zheng , Yanfei Zhong

Retrieval in 3D point clouds is a challenging task that consists in retrieving the most similar point clouds to a given query within a reference of 3D points. Current methods focus on comparing descriptors of point clouds in order to…

Computer Vision and Pattern Recognition · Computer Science 2025-05-29 Chahine-Nicolas Zede , Laurent Carrafa , Valérie Gouet-Brunet

3D reconstruction is vital for applications in autonomous driving, virtual reality, augmented reality, and the metaverse. Recent advancements such as Neural Radiance Fields(NeRF) and 3D Gaussian Splatting (3DGS) have transformed the field,…

Computer Vision and Pattern Recognition · Computer Science 2025-03-31 Zhenxiang Ma , Zhenyu Yang , Miao Tao , Yuanzhen Zhou , Zeyu He , Yuchang Zhang , Rong Fu , Hengjie Li

Point cloud 3D object detection has recently received major attention and becomes an active research topic in 3D computer vision community. However, recognizing 3D objects in LiDAR (Light Detection and Ranging) is still a challenge due to…

Computer Vision and Pattern Recognition · Computer Science 2020-10-30 Yilin Wang , Jiayi Ye

The 3D object detection capabilities in urban environments have been enormously improved by recent developments in Light Detection and Range (LiDAR) technology. This paper presents a novel framework that transforms the detection and…

Computer Vision and Pattern Recognition · Computer Science 2024-05-24 Nawfal Guefrachi , Hakim Ghazzai , Ahmad Alsharoa
‹ Prev 1 4 5 6 7 8 10 Next ›