中文
相关论文

相关论文: RoadGIE: Towards A Global-Scale Aerial Benchmark f…

200 篇论文

In this paper, we introduce a novel road marking benchmark dataset for road marking detection, addressing the limitations in the existing publicly available datasets such as lack of challenging scenarios, prominence given to lane markings,…

计算机视觉与模式识别 · 计算机科学 2022-05-06 Oshada Jayasinghe , Sahan Hemachandra , Damith Anhettigama , Shenali Kariyawasam , Ranga Rodrigo , Peshala Jayasekara

In autonomous Vehicles technology Image segmentation was a major problem in visual perception. This image segmentation process is mainly used in medical applications. Here we adopted an image segmentation process to visual perception tasks…

计算机视觉与模式识别 · 计算机科学 2020-11-17 Tirumalapudi Raviteja , Rajay Vedaraj . I. S

Streetscapes are an essential component of urban space. Their assessment is presently either limited to morphometric properties of their mass skeleton or requires labor-intensive qualitative evaluations of visually perceived qualities. This…

计算机视觉与模式识别 · 计算机科学 2026-03-18 Joan Perez , Giovanni Fusco

Ground segmentation, as the basic task of unmanned intelligent perception, provides an important support for the target detection task. Unstructured road scenes represented by open-pit mines have irregular boundary lines and uneven road…

计算机视觉与模式识别 · 计算机科学 2023-09-18 Zixuan Li , Haiying Lin , Zhangyu Wang , Huazhi Li , Miao Yu , Jie Wang

Deep learning is revolutionizing the mapping industry. Under lightweight human curation, computer has generated almost half of the roads in Thailand on OpenStreetMap (OSM) using high-resolution aerial imagery. Bing maps are displaying 125…

计算机视觉与模式识别 · 计算机科学 2019-05-07 Tao Sun , Zonglin Di , Pengyu Che , Chun Liu , Yin Wang

Pathology image segmentation is crucial in computational pathology for analyzing histological features relevant to cancer diagnosis and prognosis. However, current methods face major challenges in clinical applications due to limited…

计算机视觉与模式识别 · 计算机科学 2025-08-20 Zhixuan Chen , Junlin Hou , Liqi Lin , Yihui Wang , Yequan Bie , Xi Wang , Yanning Zhou , Ronald Cheong Kin Chan , Hao Chen

In this paper, we introduce ScenePilot-4K, a large-scale first-person dataset for safety-aware vision-language learning and evaluation in autonomous driving. Built from public online driving videos, ScenePilot-4K contains 3,847 hours of…

计算机视觉与模式识别 · 计算机科学 2026-03-31 Yujin Wang , Yutong Zheng , Wenxian Fan , Tianyi Wang , Hongqing Chu , Li Zhang , Bingzhao Gao , Daxin Tian , Jianqiang Wang , Hong Chen

We introduce OpenEarthMap, a benchmark dataset, for global high-resolution land cover mapping. OpenEarthMap consists of 2.2 million segments of 5000 aerial and satellite images covering 97 regions from 44 countries across 6 continents, with…

计算机视觉与模式识别 · 计算机科学 2022-10-20 Junshi Xia , Naoto Yokoya , Bruno Adriano , Clifford Broni-Bediako

Robust detection of AI-generated images in the wild remains challenging due to the rapid evolution of generative models and varied real-world distortions. We argue that relying on a single training regime, resolution, or backbone is…

计算机视觉与模式识别 · 计算机科学 2026-04-07 Fei Wu , Dagong Lu , Mufeng Yao , Xinlei Xu , Fengjun Guo

The fine grained classification of street trees is a crucial task for urban planning, streetscape management, and the assessment of urban ecosystem services. However, progress in this field has been hindered by the lack of large scale,…

计算机视觉与模式识别 · 计算机科学 2026-04-01 Jiapeng Li , Yingjing Huang , Fan Zhang , Yu liu

Multi-class vehicle detection from airborne imagery with orientation estimation is an important task in the near and remote vision domains with applications in traffic monitoring and disaster management. In the last decade, we have…

计算机视觉与模式识别 · 计算机科学 2020-11-25 Seyed Majid Azimi , Reza Bahmanyar , Corenin Henry , Franz Kurz

World models aim to simulate environments and enable effective agent behavior. However, modeling real-world environments presents unique challenges as they dynamically change across both space and, crucially, time. To capture these composed…

Recently, road scene-graph representations used in conjunction with graph learning techniques have been shown to outperform state-of-the-art deep learning techniques in tasks including action classification, risk assessment, and collision…

计算机视觉与模式识别 · 计算机科学 2022-01-03 Arnav Vaibhav Malawade , Shih-Yuan Yu , Brandon Hsu , Harsimrat Kaeley , Anurag Karra , Mohammad Abdullah Al Faruque

Recent advances in interactive 3D segmentation from 2D images have demonstrated impressive performance. However, current models typically require extensive scene-specific training to accurately reconstruct and segment objects, which limits…

计算机视觉与模式识别 · 计算机科学 2025-03-18 Yansong Guo , Jie Hu , Yansong Qu , Liujuan Cao

Research on damage detection of road surfaces has been an active area of re-search, but most studies have focused so far on the detection of the presence of damages. However, in real-world scenarios, road managers need to clearly understand…

计算机视觉与模式识别 · 计算机科学 2022-01-21 A. A. Angulo , J. A. Vega-Fernández , L. M. Aguilar-Lobo , S. Natraj , G Ochoa-Ruiz

Robots operating in unstructured environments often require accurate and consistent object-level representations. This typically requires segmenting individual objects from the robot's surroundings. While recent large models such as Segment…

机器人学 · 计算机科学 2025-04-07 Haozhan Tang , Tianyi Zhang , Oliver Kroemer , Matthew Johnson-Roberson , Weiming Zhi

Interactive point-based image editing serves as a controllable editor, enabling precise and flexible manipulation of image content. However, most drag-based methods operate primarily on the 2D pixel plane with limited use of 3D cues. As a…

计算机视觉与模式识别 · 计算机科学 2026-02-23 Xinyu Pu , Hongsong Wang , Jie Gui , Pan Zhou

Existing Earth Vision datasets are either suitable for semantic segmentation or object detection. In this work, we introduce the first benchmark dataset for instance segmentation in aerial imagery that combines instance-level object…

计算机视觉与模式识别 · 计算机科学 2019-08-29 Syed Waqas Zamir , Aditya Arora , Akshita Gupta , Salman Khan , Guolei Sun , Fahad Shahbaz Khan , Fan Zhu , Ling Shao , Gui-Song Xia , Xiang Bai

Urban segmentation and lane detection are two important tasks for traffic scene perception. Accuracy and fast inference speed of visual perception are crucial for autonomous driving safety. Fine and complex geometric objects are the most…

计算机视觉与模式识别 · 计算机科学 2024-04-04 Yaxin Feng , Yuan Lan , Luchan Zhang , Guoqing Liu , Yang Xiang

Accurate flood detection from visual data is a critical step toward improving disaster response and risk assessment, yet datasets for flood segmentation remain scarce due to the challenges of collecting and annotating large-scale imagery.…

计算机视觉与模式识别 · 计算机科学 2026-03-23 Georgios Simantiris , Konstantinos Bacharidis , Apostolos Papanikolaou , Petros Giannakakis , Costas Panagiotakis