English
Related papers

Related papers: THUD++: Large-Scale Dynamic Indoor Scene Dataset a…

200 papers

This paper presents Ev-Layout, a novel large-scale event-based multi-modal dataset designed for indoor layout estimation and tracking. Ev-Layout makes key contributions to the community by: Utilizing a hybrid data collection platform (with…

Graphics · Computer Science 2025-03-12 Xucheng Guo , Yiran Shen , Xiaofang Xiao , Yuanfeng Zhou , Lin Wang

With the acceleration of urbanization and the growth of transportation demands, the safety of vulnerable road users (VRUs, such as pedestrians and cyclists) in mixed traffic flows has become increasingly prominent, necessitating…

Computer Vision and Pattern Recognition · Computer Science 2026-04-30 Zhangcun Yan , Jianqiang Li , Peng Hang , Jian Sun

Advances in neural fields are enabling high-fidelity capture of the shape and appearance of dynamic 3D scenes. However, their capabilities lag behind those offered by conventional representations such as 2D videos because of algorithmic…

Computer Vision and Pattern Recognition · Computer Science 2024-12-06 Cheng-You Lu , Peisen Zhou , Angela Xing , Chandradeep Pokhariya , Arnab Dey , Ishaan Shah , Rugved Mavidipalli , Dylan Hu , Andrew Comport , Kefan Chen , Srinath Sridhar

Dynamic scene understanding is the ability of a computer system to interpret and make sense of the visual information present in a video of a real-world scene. In this thesis, we present a series of frameworks for dynamic scene…

Computer Vision and Pattern Recognition · Computer Science 2023-12-14 Salman Khan

Generally, crowd datasets can be collected or generated from real or synthetic sources. Real data is generated by using infrastructure-based sensors (such as static cameras or other sensors). The use of simulation tools can significantly…

Computer Vision and Pattern Recognition · Computer Science 2023-04-27 Paweł Foszner , Agnieszka Szczęsna , Luca Ciampi , Nicola Messina , Adam Cygan , Bartosz Bizoń , Michał Cogiel , Dominik Golba , Elżbieta Macioszek , Michał Staniszewski

We present UbiSLAM, an innovative solution for real-time mapping and localization in dynamic indoor environments. By deploying a network of fixed RGB-D cameras strategically throughout the workspace, UbiSLAM addresses limitations commonly…

Robotics · Computer Science 2026-05-19 Halim Djerroud , Nico Steyn , Olivier Rabreau , Patrick Bonnin , Abderraouf Benali

In this data article, we introduce the Multi-Modal Event-based Vehicle Detection and Tracking (MEVDT) dataset. This dataset provides a synchronized stream of event data and grayscale images of traffic scenes, captured using the Dynamic and…

Computer Vision and Pattern Recognition · Computer Science 2024-07-31 Zaid A. El Shair , Samir A. Rawashdeh

Humans drive in a holistic fashion which entails, in particular, understanding dynamic road events and their evolution. Injecting these capabilities in autonomous vehicles can thus take situational awareness and decision making closer to…

We present MOSU, a novel autonomous long-range navigation system that enhances global navigation for mobile robots through multimodal perception and on-road scene understanding. MOSU addresses the outdoor robot navigation challenge by…

Robotics · Computer Science 2025-07-08 Jing Liang , Kasun Weerakoon , Daeun Song , Senthurbavan Kirubaharan , Xuesu Xiao , Dinesh Manocha

Robotic task planning in real-world environments requires not only object recognition but also a nuanced understanding of spatial relationships between objects. We present a spatial-relationship-aware dataset of nearly 1,000 robot-acquired…

Robotics · Computer Science 2025-06-17 Peng Wang , Minh Huy Pham , Zhihao Guo , Wei Zhou

Analyzing scenes thoroughly is crucial for mobile robots acting in different environments. Semantic segmentation can enhance various subsequent tasks, such as (semantically assisted) person perception, (semantic) free space detection,…

Computer Vision and Pattern Recognition · Computer Science 2021-04-08 Daniel Seichter , Mona Köhler , Benjamin Lewandowski , Tim Wengefeld , Horst-Michael Gross

The creation of large, diverse, high-quality robot manipulation datasets is an important stepping stone on the path toward more capable and robust robotic manipulation policies. However, creating such datasets is challenging: collecting…

Robotics · Computer Science 2025-04-23 Alexander Khazatsky , Karl Pertsch , Suraj Nair , Ashwin Balakrishna , Sudeep Dasari , Siddharth Karamcheti , Soroush Nasiriany , Mohan Kumar Srirama , Lawrence Yunliang Chen , Kirsty Ellis , Peter David Fagan , Joey Hejna , Masha Itkina , Marion Lepert , Yecheng Jason Ma , Patrick Tree Miller , Jimmy Wu , Suneel Belkhale , Shivin Dass , Huy Ha , Arhan Jain , Abraham Lee , Youngwoon Lee , Marius Memmel , Sungjae Park , Ilija Radosavovic , Kaiyuan Wang , Albert Zhan , Kevin Black , Cheng Chi , Kyle Beltran Hatch , Shan Lin , Jingpei Lu , Jean Mercat , Abdul Rehman , Pannag R Sanketi , Archit Sharma , Cody Simpson , Quan Vuong , Homer Rich Walke , Blake Wulfe , Ted Xiao , Jonathan Heewon Yang , Arefeh Yavary , Tony Z. Zhao , Christopher Agia , Rohan Baijal , Mateo Guaman Castro , Daphne Chen , Qiuyu Chen , Trinity Chung , Jaimyn Drake , Ethan Paul Foster , Jensen Gao , Vitor Guizilini , David Antonio Herrera , Minho Heo , Kyle Hsu , Jiaheng Hu , Muhammad Zubair Irshad , Donovon Jackson , Charlotte Le , Yunshuang Li , Kevin Lin , Roy Lin , Zehan Ma , Abhiram Maddukuri , Suvir Mirchandani , Daniel Morton , Tony Nguyen , Abigail O'Neill , Rosario Scalise , Derick Seale , Victor Son , Stephen Tian , Emi Tran , Andrew E. Wang , Yilin Wu , Annie Xie , Jingyun Yang , Patrick Yin , Yunchu Zhang , Osbert Bastani , Glen Berseth , Jeannette Bohg , Ken Goldberg , Abhinav Gupta , Abhishek Gupta , Dinesh Jayaraman , Joseph J Lim , Jitendra Malik , Roberto Martín-Martín , Subramanian Ramamoorthy , Dorsa Sadigh , Shuran Song , Jiajun Wu , Michael C. Yip , Yuke Zhu , Thomas Kollar , Sergey Levine , Chelsea Finn

Layout estimation and 3D object detection are two fundamental tasks in indoor scene understanding. When combined, they enable the creation of a compact yet semantically rich spatial representation of a scene. Existing approaches typically…

Computer Vision and Pattern Recognition · Computer Science 2025-09-29 Anton Konushin , Nikita Drozdov , Bulat Gabdullin , Alexey Zakharov , Anna Vorontsova , Danila Rukhovich , Maksim Kolodiazhnyi

Datasets have gained an enormous amount of popularity in the computer vision community, from training and evaluation of Deep Learning-based methods to benchmarking Simultaneous Localization and Mapping (SLAM). Without a doubt, synthetic…

Computer Vision and Pattern Recognition · Computer Science 2018-09-05 Wenbin Li , Sajad Saeedi , John McCormac , Ronald Clark , Dimos Tzoumanikas , Qing Ye , Yuzhong Huang , Rui Tang , Stefan Leutenegger

Urbanization is advancing rapidly, covering less than 2% of Earth's surface yet profoundly influencing global environments and experiencing disproportionate impacts from extreme weather events. Effective urban management and planning…

Visual Object Tracking (VOT) is a fundamental task with widespread applications in autonomous navigation, surveillance, and maritime robotics. Despite significant advances in generic object tracking, maritime environments continue to…

Computer Vision and Pattern Recognition · Computer Science 2025-06-04 Ahsan Baidar Bakht , Muhayy Ud Din , Sajid Javed , Irfan Hussain

Perception is a cornerstone of autonomous driving, enabling vehicles to understand their surroundings and make safe, reliable decisions. Developing robust perception algorithms requires large-scale, high-quality datasets that cover diverse…

Computer Vision and Pattern Recognition · Computer Science 2026-01-30 Dominik Rößle , Xujun Xie , Adithya Mohan , Venkatesh Thirugnana Sambandham , Daniel Cremers , Torsten Schön

The real-world deployment of fully autonomous mobile robots depends on a robust SLAM (Simultaneous Localization and Mapping) system, capable of handling dynamic environments, where objects are moving in front of the robot, and changing…

With the recent rise of Large Language Models (LLMs), Vision-Language Models (VLMs), and other general foundation models, there is growing potential for multimodal, multi-task embodied agents that can operate in diverse environments given…

Robotics · Computer Science 2024-11-07 Haochen Zhang , Nader Zantout , Pujith Kachana , Zongyuan Wu , Ji Zhang , Wenshan Wang

We present Urban-ImageNet, a large-scale multi-modal dataset and evaluation benchmark for urban space perception from user-generated social media imagery. The corpus contains over 2 Million public social media images and paired textual…

Computer Vision and Pattern Recognition · Computer Science 2026-05-12 Yiwei Ou , Chung Ching Cheung , Jun Yang Ang , Xiaobin Ren , Ronggui Sun , Guansong Gao , Kaiqi Zhao , Manfredo Manfredini