中文
相关论文

相关论文: G2D: from GTA to Data

200 篇论文

Pedestrian detection through Computer Vision is a building block for a multitude of applications. Recently, there was an increasing interest in Convolutional Neural Network-based architectures for the execution of such a task. One of these…

计算机视觉与模式识别 · 计算机科学 2020-09-22 Luca Ciampi , Nicola Messina , Fabrizio Falchi , Claudio Gennaro , Giuseppe Amato

In computer vision, the development of robust algorithms capable of generalizing effectively in real-world scenarios more and more often requires large-scale datasets collected under diverse environmental conditions. However, acquiring such…

计算机视觉与模式识别 · 计算机科学 2025-02-19 Matteo Scucchia , Matteo Ferrara , Davide Maltoni

Cloud gaming has gained popularity as it provides high-quality gaming experiences on thin hardware, such as phones and tablets. Transmitting gameplay frames at high resolutions and ultra-low latency is the key to guaranteeing players'…

数据库 · 计算机科学 2025-08-11 Hongqin Lei , Haowei Tang , Zhe Zhang

Image- and video-based 3D human recovery (i.e., pose and shape estimation) have achieved substantial progress. However, due to the prohibitive cost of motion capture, existing datasets are often limited in scale and diversity. In this work,…

计算机视觉与模式识别 · 计算机科学 2024-09-11 Zhongang Cai , Mingyuan Zhang , Jiawei Ren , Chen Wei , Daxuan Ren , Zhengyu Lin , Haiyu Zhao , Lei Yang , Chen Change Loy , Ziwei Liu

High Dynamic Range (HDR) content (i.e., images and videos) has a broad range of applications. However, capturing HDR content from real-world scenes is expensive and time-consuming. Therefore, the challenging task of reconstructing visually…

计算机视觉与模式识别 · 计算机科学 2024-03-28 Hrishav Bakul Barua , Kalin Stefanov , KokSheik Wong , Abhinav Dhall , Ganesh Krishnasamy

We propose to build realistic virtual worlds, called 360RVW, for large urban environments directly from 360{\deg} videos. We provide an interface for interactive exploration, where users can freely navigate via their own avatars. 360{\deg}…

多媒体 · 计算机科学 2025-10-14 Mizuki Takenawa , Naoki Sugimoto , Leslie Wöhler , Satoshi Ikehata , Kiyoharu Aizawa

Recent advancements in video anomaly detection (VAD) have enabled identification of various criminal activities in surveillance videos, but detecting fatal incidents such as shootings and stabbings remains difficult due to their rarity and…

计算机视觉与模式识别 · 计算机科学 2025-09-11 Seongho Kim , Sejong Ryu , Hyoukjun You , Je Hyeong Hong

We introduce the Precise Synthetic Image and LiDAR (PreSIL) dataset for autonomous vehicle perception. Grand Theft Auto V (GTA V), a commercial video game, has a large detailed world with realistic graphics, which provides a diverse data…

计算机视觉与模式识别 · 计算机科学 2019-05-08 Braden Hurl , Krzysztof Czarnecki , Steven Waslander

People's transportation choices reflect complex trade-offs shaped by personal preferences, social norms, and technology acceptance. Predicting such behavior at scale is a critical challenge with major implications for urban planning and…

人机交互 · 计算机科学 2026-01-28 Simon Lämmer , Mark Colley , Patrick Ebel

Satellites are capable of capturing high-resolution videos. It makes vehicle perception from satellite become possible. Compared to street surveillance, drive recorder or other equipments, satellite videos provide a much broader city-scale…

计算机视觉与模式识别 · 计算机科学 2024-03-27 Bin Zhao , Pengfei Han , Xuelong Li

Recent developments in generative models and large-scale datasets have substantially advanced 3D world generation, facilitating a broad range of domains including spatial intelligence, embodied intelligence, and autonomous driving. While…

计算机视觉与模式识别 · 计算机科学 2026-05-14 Hanxin Zhu , Cong Wang , Peiyan Tu , Jiayi Luo , Tianyu He , Xin Jin , Zhibo Chen

The vision-based geo-localization technology for UAV, serving as a secondary source of GPS information in addition to the global navigation satellite systems (GNSS), can still operate independently in the GPS-denied environment. Recent deep…

计算机视觉与模式识别 · 计算机科学 2024-12-13 Yuxiang Ji , Boyong He , Zhuoyue Tan , Liaoni Wu

Large-scale video generation models have demonstrated high visual realism in diverse contexts, spurring interest in their potential as general-purpose world simulators. Existing benchmarks focus on individual subjects rather than scenes…

计算机视觉与模式识别 · 计算机科学 2025-10-24 Aaron Appelle , Jerome P. Lynch

Vehicle-to-Vehicle (V2V) cooperative perception has great potential to enhance autonomous driving performance by overcoming perception limitations in complex adverse traffic scenarios (CATS). Meanwhile, data serves as the fundamental…

Driving datasets accelerate the development of intelligent driving and related computer vision technologies, while substantial and detailed annotations serve as fuels and powers to boost the efficacy of such datasets to improve…

机器学习 · 计算机科学 2019-06-04 Zhengping Che , Guangyu Li , Tracy Li , Bo Jiang , Xuefeng Shi , Xinsheng Zhang , Ying Lu , Guobin Wu , Yan Liu , Jieping Ye

When humans play virtual racing games, they use visual environmental information on the game screen to understand the rules within the environments. In contrast, a state-of-the-art realistic racing game AI agent that outperforms human…

人工智能 · 计算机科学 2021-11-15 Ryuji Imamura , Takuma Seno , Kenta Kawamoto , Michael Spranger

The lack of large-scale datasets has been impeding the advance of deep learning approaches to the problem of F-formation detection. Moreover, most research works on this problem rely on input sensor signals of object location and…

计算机视觉与模式识别 · 计算机科学 2022-11-22 Giang Hoang , Tuan Nguyen Dinh , Tung Cao Hoang , Son Le Duy , Keisuke Hihara , Yumeka Utada , Akihiko Torii , Naoki Izumi , Long Tran Quoc

Simulation is a prospective method for generating diverse and realistic traffic scenarios to aid in the development of driving decision-making systems. However, existing simulators often fall short in diverse scenarios or interactive…

机器学习 · 计算机科学 2024-05-21 Yueyuan Li , Songan Zhang , Mingyang Jiang , Xingyuan Chen , Yeqiang Qian , Chunxiang Wang , Ming Yang

Corner cases are crucial for training and validating autonomous driving systems, yet collecting them from the real world is often costly and hazardous. Editing objects within captured sensor data offers an effective alternative for…

计算机视觉与模式识别 · 计算机科学 2025-08-29 Jiusi Li , Jackson Jiang , Jinyu Miao , Miao Long , Tuopu Wen , Peijin Jia , Shengxiang Liu , Chunlei Yu , Maolin Liu , Yuzhan Cai , Kun Jiang , Mengmeng Yang , Diange Yang

The current autonomous stack is well modularized and consists of perception, decision making and control in a handcrafted framework. With the advances in artificial intelligence (AI) and computing resources, researchers have been pushing…

机器人学 · 计算机科学 2024-04-19 Satya R. Jaladi , Zhimin Chen , Narahari R. Malayanur , Raja M. Macherla , Bing Li
‹ 上一页 1 2 3 10 下一页 ›