中文
相关论文

相关论文: SkyScenes: A Synthetic Dataset for Aerial Scene Un…

200 篇论文

Dynamic scene understanding is the ability of a computer system to interpret and make sense of the visual information present in a video of a real-world scene. In this thesis, we present a series of frameworks for dynamic scene…

计算机视觉与模式识别 · 计算机科学 2023-12-14 Salman Khan

The success of deep learning in computer vision is based on availability of large annotated datasets. To lower the need for hand labeled images, virtually rendered 3D worlds have recently gained popularity. Creating realistic 3D content is…

计算机视觉与模式识别 · 计算机科学 2017-08-07 Hassan Abu Alhaija , Siva Karthik Mustikovela , Lars Mescheder , Andreas Geiger , Carsten Rother

To advance research in learning-based defogging algorithms, various synthetic fog datasets have been developed. However, existing datasets created using the Atmospheric Scattering Model (ASM) or real-time rendering engines often struggle to…

计算机视觉与模式识别 · 计算机科学 2024-03-27 Yiming Xie , Henglu Wei , Zhenyi Liu , Xiaoyu Wang , Xiangyang Ji

Creating a diverse and comprehensive dataset of hand gestures for dynamic human-machine interfaces in the automotive domain can be challenging and time-consuming. To overcome this challenge, we propose using synthetic gesture datasets…

计算机视觉与模式识别 · 计算机科学 2024-08-05 Amr Gomaa , Robin Zitt , Guillermo Reyes , Antonio Krüger

Classical and more recently deep computer vision methods are optimized for visible spectrum images, commonly encoded in grayscale or RGB colorspaces acquired from smartphones or cameras. A more uncommon source of images exploited in the…

计算机视觉与模式识别 · 计算机科学 2020-01-29 Caio C. V. da Silva , Keiller Nogueira , Hugo N. Oliveira , Jefersson A. dos Santos

Human vision is capable of transforming two-dimensional observations into an egocentric three-dimensional scene understanding, which underpins the ability to translate complex scenes and exhibit adaptive behaviors. This capability, however,…

计算机视觉与模式识别 · 计算机科学 2025-09-26 Pei Liu , Hongliang Lu , Haichao Liu , Haipeng Liu , Xin Liu , Ruoyu Yao , Shengbo Eben Li , Jun Ma

The visual entities in cross-view images exhibit drastic domain changes due to the difference in viewpoints each set of images is captured from. Existing state-of-the-art methods address the problem by learning view-invariant descriptors…

计算机视觉与模式识别 · 计算机科学 2019-08-12 Krishna Regmi , Mubarak Shah

We present Eyecandies, a novel synthetic dataset for unsupervised anomaly detection and localization. Photo-realistic images of procedurally generated candies are rendered in a controlled environment under multiple lightning conditions,…

计算机视觉与模式识别 · 计算机科学 2022-10-11 Luca Bonfiglioli , Marco Toschi , Davide Silvestri , Nicola Fioraio , Daniele De Gregorio

With recent developments in Embodied Artificial Intelligence (EAI) research, there has been a growing demand for high-quality, large-scale interactive scene generation. While prior methods in scene synthesis have prioritized the naturalness…

计算机视觉与模式识别 · 计算机科学 2024-07-11 Yandan Yang , Baoxiong Jia , Peiyuan Zhi , Siyuan Huang

Understanding road scenes for visual perception remains crucial for intelligent self-driving cars. In particular, it is desirable to detect unexpected small road hazards reliably in real-time, especially under varying adverse conditions…

计算机视觉与模式识别 · 计算机科学 2025-12-30 Jongoh Jeong , Taek-Jin Song , Jong-Hwan Kim , Kuk-Jin Yoon

The development of aerial holistic scene understanding algorithms is hindered by the scarcity of comprehensive datasets that enable both semantic and geometric reconstruction. While synthetic datasets offer an alternative, existing options…

计算机视觉与模式识别 · 计算机科学 2025-08-01 Radu Beche , Sergiu Nedevschi

Indoor scene augmentation has become an emerging topic in the field of computer vision and graphics with applications in augmented and virtual reality. However, current state-of-the-art systems using deep neural networks require large…

计算机视觉与模式识别 · 计算机科学 2021-03-30 Mohammad Keshavarzi , Flaviano Christian Reyes , Ritika Shrivastava , Oladapo Afolabi , Luisa Caldas , Allen Y. Yang

In recent years, street view imagery has grown to become one of the most important sources of geospatial data collection and urban analytics, which facilitates generating meaningful insights and assisting in decision-making. Synthesizing a…

计算机视觉与模式识别 · 计算机科学 2025-10-01 Khawlah Bajbaa , Muhammad Usman , Saeed Anwar , Ibrahim Radwan , Abdul Bais

Cinematic video production requires control over scene-subject composition and camera movement, but live-action shooting remains costly due to the need for constructing physical sets. To address this, we introduce the task of cinematic…

计算机视觉与模式识别 · 计算机科学 2026-02-09 Kaiyi Huang , Yukun Huang , Yu Li , Jianhong Bai , Xintao Wang , Zinan Lin , Xuefei Ning , Jiwen Yu , Pengfei Wan , Yu Wang , Xihui Liu

Current computer vision tasks based on deep learning require a huge amount of data with annotations for model training or testing, especially in some dense estimation tasks, such as optical flow segmentation and depth estimation. In…

计算机视觉与模式识别 · 计算机科学 2022-06-29 Xiangtong Wang , Binbin Liang , Menglong Yang , Wei Li

Interest in robotics for forest management is growing, but perception in complex, natural environments remains a significant hurdle. Conditions such as heavy occlusion, variable lighting, and dense vegetation pose challenges to automated…

We present a challenging dataset, the TartanAir, for robot navigation tasks and more. The data is collected in photo-realistic simulation environments with the presence of moving objects, changing light and various weather conditions. By…

机器人学 · 计算机科学 2020-08-11 Wenshan Wang , Delong Zhu , Xiangwei Wang , Yaoyu Hu , Yuheng Qiu , Chen Wang , Yafei Hu , Ashish Kapoor , Sebastian Scherer

Autonomous driving system development is critically dependent on the ability to replay complex and diverse traffic scenarios in simulation. In such scenarios, the ability to accurately simulate the vehicle sensors such as cameras, lidar or…

计算机视觉与模式识别 · 计算机科学 2020-06-26 Zhenpei Yang , Yuning Chai , Dragomir Anguelov , Yin Zhou , Pei Sun , Dumitru Erhan , Sean Rafferty , Henrik Kretzschmar

Semantic segmentation of drone images is critical for various aerial vision tasks as it provides essential semantic details to understand scenes on the ground. Ensuring high accuracy of semantic segmentation models for drones requires…

计算机视觉与模式识别 · 计算机科学 2025-04-01 Wenxiao Cai , Ke Jin , Jinyan Hou , Cong Guo , Letian Wu , Wankou Yang

Humans possess the cognitive ability to comprehend scenes in a compositional manner. To empower AI systems with similar capabilities, object-centric learning aims to acquire representations of individual objects from visual scenes without…

计算机视觉与模式识别 · 计算机科学 2023-09-07 Yinxuan Huang , Tonglin Chen , Zhimeng Shen , Jinghao Huang , Bin Li , Xiangyang Xue