中文
相关论文

相关论文: SkyScenes: A Synthetic Dataset for Aerial Scene Un…

200 篇论文

Semantic segmentation has been one of the leading research interests in computer vision recently. It serves as a perception foundation for many fields, such as robotics and autonomous driving. The fast development of semantic segmentation…

计算机视觉与模式识别 · 计算机科学 2020-05-19 Ye Lyu , George Vosselman , Guisong Xia , Alper Yilmaz , Michael Ying Yang

Simulation is crucial for developing and evaluating autonomous vehicle (AV) systems. Recent literature builds on a new generation of generative models to synthesize highly realistic images for full-stack simulation. However, purely…

计算机视觉与模式识别 · 计算机科学 2025-06-25 Zehao Zhu , Yuliang Zou , Chiyu Max Jiang , Bo Sun , Vincent Casser , Xiukun Huang , Jiahao Wang , Zhenpei Yang , Ruiqi Gao , Leonidas Guibas , Mingxing Tan , Dragomir Anguelov

Unmanned Aerial Vehicle (UAV) has gained significant traction in the recent years, particularly the context of surveillance. However, video datasets that capture violent and non-violent human activity from aerial point-of-view is scarce. To…

计算机视觉与模式识别 · 计算机科学 2022-08-16 Mahieyin Rahmun , Tonmoay Deb , Shahriar Ali Bijoy , Mayamin Hamid Raha

Diffusion models are advancing autonomous driving by enabling realistic data synthesis, predictive end-to-end planning, and closed-loop simulation, with a primary focus on temporally consistent generation. However, large-scale 3D scene…

计算机视觉与模式识别 · 计算机科学 2025-12-09 Yu Yang , Alan Liang , Jianbiao Mei , Yukai Ma , Yong Liu , Gim Hee Lee

This paper introduces SynTraC, the first public image-based traffic signal control dataset, aimed at bridging the gap between simulated environments and real-world traffic management challenges. Unlike traditional datasets for traffic…

人工智能 · 计算机科学 2024-08-20 Tiejin Chen , Prithvi Shirke , Bharatesh Chakravarthi , Arpitsinh Vaghela , Longchao Da , Duo Lu , Yezhou Yang , Hua Wei

Recent advances in Computer Vision and Deep Learning have enabled astonishing results in a variety of fields and applications. Motivated by this success, the SkyCam Dataset aims to enable image-based Deep Learning solutions for short-term,…

计算机视觉与模式识别 · 计算机科学 2021-05-10 Evangelos Ntavelis , Jan Remund , Philipp Schmid

In this work we consider UAVs as cooperative agents supporting human users in their operations. In this context, the 3D localisation of the UAV assistant is an important task that can facilitate the exchange of spatial information between…

计算机视觉与模式识别 · 计算机科学 2020-08-24 Georgios Albanis , Nikolaos Zioulis , Anastasios Dimou , Dimitrios Zarpalas , Petros Daras

Video image datasets are playing an essential role in design and evaluation of traffic vision algorithms. Nevertheless, a longstanding inconvenience concerning image datasets is that manually collecting and annotating large-scale…

计算机视觉与模式识别 · 计算机科学 2017-12-25 Xuan Li , Kunfeng Wang , Yonglin Tian , Lan Yan , Fei-Yue Wang

Vision-and-language navigation (VLN) aims to develop agents capable of navigating in realistic environments. While recent cross-modal training approaches have significantly improved navigation performance in both indoor and outdoor…

计算机视觉与模式识别 · 计算机科学 2025-08-05 Jungdae Lee , Taiki Miyanishi , Shuhei Kurita , Koya Sakamoto , Daichi Azuma , Yutaka Matsuo , Nakamasa Inoue

Autonomous trucking is a promising technology that can greatly impact modern logistics and the environment. Ensuring its safety on public roads is one of the main duties that requires an accurate perception of the environment. To achieve…

Learning on synthetic data and transferring the resulting properties to their real counterparts is an important challenge for reducing costs and increasing safety in machine learning. In this work, we focus on autoencoder architectures and…

计算机视觉与模式识别 · 计算机科学 2022-04-04 Steve Dias Da Cruz , Bertram Taetz , Thomas Stifter , Didier Stricker

Unmanned Aerial Vehicles (UAVs), equipped with camera sensors can facilitate enhanced situational awareness for many emergency response and disaster management applications since they are capable of operating in remote and difficult to…

计算机视觉与模式识别 · 计算机科学 2019-06-21 Christos Kyrkou , Theocharis Theocharides

Deep neural networks (DNN) which are employed in perception systems for autonomous driving require a huge amount of data to train on, as they must reliably achieve high performance in all kinds of situations. However, these DNN are usually…

机器人学 · 计算机科学 2023-08-01 Daniel Bogdoll , Svenja Uhlemeyer , Kamil Kowol , J. Marius Zöllner

We present a challenging dataset, ChangeSim, aimed at online scene change detection (SCD) and more. The data is collected in photo-realistic simulation environments with the presence of environmental non-targeted variations, such as air…

计算机视觉与模式识别 · 计算机科学 2021-07-23 Jin-Man Park , Jae-Hyuk Jang , Sahng-Min Yoo , Sun-Kyung Lee , Ue-Hwan Kim , Jong-Hwan Kim

Scene understanding is an active research area. Commercial depth sensors, such as Kinect, have enabled the release of several RGB-D datasets over the past few years which spawned novel methods in 3D scene understanding. More recently with…

计算机视觉与模式识别 · 计算机科学 2022-01-13 Gilad Baruch , Zhuoyuan Chen , Afshin Dehghan , Tal Dimry , Yuri Feigin , Peter Fu , Thomas Gebauer , Brandon Joffe , Daniel Kurz , Arik Schwartz , Elad Shulman

Unmanned Aerial Vehicles (UAVs) rely on satellite systems for stable positioning. However, due to limited satellite coverage or communication disruptions, UAVs may lose signals from satellite-based positioning systems. In such situations,…

计算机视觉与模式识别 · 计算机科学 2023-08-14 Ming Dai , Enhui Zheng , Zhenhua Feng , Jiedong Zhuang , Wankou Yang

This paper presents a simulation workflow for generating synthetic LiDAR datasets to support autonomous vehicle perception, robotics research, and sensor security analysis. Leveraging the CoppeliaSim simulation environment and its Python…

机器人学 · 计算机科学 2025-06-24 Abhishek Phadke , Shakib Mahmud Dipto , Pratip Rana

Remote sensing through unmanned aerial systems (UAS) has been increasing in forestry in recent years, along with using machine learning for data processing. Deep learning architectures, extensively applied in natural language and image…

计算机视觉与模式识别 · 计算机科学 2024-04-19 Francisco Raverta Capua , Juan Schandin , Pablo De Cristóforis

This study aims to investigate the challenge of insufficient three-dimensional context in synthetic datasets for scene text rendering. Although recent advances in diffusion models and related techniques have improved certain aspects of…

计算机视觉与模式识别 · 计算机科学 2025-05-27 Li-Syun Hsiung , Jun-Kai Tu , Kuan-Wu Chu , Yu-Hsuan Chiu , Yan-Tsung Peng , Sheng-Luen Chung , Gee-Sern Jison Hsu

Robust and realistic rendering for large-scale road scenes is essential in autonomous driving simulation. Recently, 3D Gaussian Splatting (3D-GS) has made groundbreaking progress in neural rendering, but the general fidelity of large-scale…

计算机视觉与模式识别 · 计算机科学 2024-08-28 Saining Zhang , Baijun Ye , Xiaoxue Chen , Yuantao Chen , Zongzheng Zhang , Cheng Peng , Yongliang Shi , Hao Zhao