English
Related papers

Related papers: YouTube AV 50K: An Annotated Corpus for Comments i…

200 papers

Crash data of autonomous vehicles (AV) or vehicles equipped with advanced driver assistance systems (ADAS) are the key information to understand the crash nature and to enhance the automation systems. However, most of the existing crash…

Robotics · Computer Science 2023-03-24 Ou Zheng , Mohamed Abdel-Aty , Zijin Wang , Shengxuan Ding , Dongdong Wang , Yuxuan Huang

Existing video-language models can generate factual descriptions of road events but lack control over how these events are expressed: their tone, urgency, or style. This limits deployment in communication-critical settings where the…

Computer Vision and Pattern Recognition · Computer Science 2026-05-21 Chirag Parikh , Siddhi Pravin Lipare , Ravi Kiran Sarvadevabhatla

The end-to-end learning ability of self-driving vehicles has achieved significant milestones over the last decade owing to rapid advances in deep learning and computer vision algorithms. However, as autonomous driving technology is a…

Computer Vision and Pattern Recognition · Computer Science 2023-07-21 Shahin Atakishiyev , Mohammad Salameh , Housam Babiker , Randy Goebel

YouTube, a widely popular online platform, has transformed the dynamics of con-tent consumption and interaction for users worldwide. With its extensive range of content crea-tors and viewers, YouTube serves as a hub for video sharing,…

Social and Information Networks · Computer Science 2023-11-13 Shadi Shajari , Mustafa Alassad , Nitin Agarwal

Multi-modal retrieval is an important problem for many applications, such as recommendation and search. Current benchmarks and even datasets are often manually constructed and consist of mostly clean samples where all modalities are…

Computer Vision and Pattern Recognition · Computer Science 2022-10-21 Laura Hanu , James Thewlis , Yuki M. Asano , Christian Rupprecht

This survey offers a comprehensive examination of collaborative perception datasets in the context of Vehicle-to-Infrastructure (V2I), Vehicle-to-Vehicle (V2V), and Vehicle-to-Everything (V2X). It highlights the latest developments in…

Computer Vision and Pattern Recognition · Computer Science 2024-04-23 Melih Yazgan , Mythra Varun Akkanapragada , J. Marius Zoellner

We propose a novel framework for predicting the factuality of reporting of news media outlets by studying the user attention cycles in their YouTube channels. In particular, we design a rich set of features derived from the temporal…

Computation and Language · Computer Science 2021-08-31 Krasimira Bozhanova , Yoan Dinkov , Ivan Koychev , Maria Castaldo , Tommaso Venturini , Preslav Nakov

Automated Vehicles (AVs) promise significant advances in transportation. Critical to these improvements is understanding AVs' longitudinal behavior, relying heavily on real-world trajectory data. Existing open-source trajectory datasets of…

Robotics · Computer Science 2025-04-29 Hang Zhou , Ke Ma , Shixiao Liang , Xiaopeng Li , Xiaobo Qu

To ensure safe operation of autonomous vehicles in complex urban environments, complete perception of the environment is necessary. However, due to environmental conditions, sensor limitations, and occlusions, this is not always possible…

Computer Vision and Pattern Recognition · Computer Science 2024-05-28 Sven Teufel , Jörg Gamerdinger , Jan-Patrick Kirchner , Georg Volk , Oliver Bringmann

Video watching had emerged as one of the most frequent media activities on the Internet. Yet, little is known about how users watch online video. Using two distinct YouTube datasets, a set of random YouTube videos crawled from the Web and a…

Human-Computer Interaction · Computer Science 2017-05-18 Minsu Park , Mor Naaman , Jonah Berger

Videos can evoke a range of affective responses in viewers. The ability to predict evoked affect from a video, before viewers watch the video, can help in content creation and video recommendation. We introduce the Evoked Expressions from…

Computer Vision and Pattern Recognition · Computer Science 2021-02-23 Jennifer J. Sun , Ting Liu , Alan S. Cowen , Florian Schroff , Hartwig Adam , Gautam Prasad

Human affect recognition has been a significant topic in psychophysics and computer vision. However, the currently published datasets have many limitations. For example, most datasets contain frames that contain only information about…

Computer Vision and Pattern Recognition · Computer Science 2023-09-18 Zhihang Ren , Jefferson Ortega , Yifan Wang , Zhimin Chen , Yunhui Guo , Stella X. Yu , David Whitney

Violent threats remain a significant problem across social media platforms. Useful, high-quality data facilitates research into the understanding and detection of malicious content, including violence. In this paper, we introduce a…

Video sharing sites, such as YouTube, use video responses to enhance the social interactions among their users. The video response feature allows users to interact and converse through video, by creating a video sequence that begins with an…

Short-video platforms show an increasing impact on people's daily lives nowadays, with billions of active users spending plenty of time each day. The interactions between users and online platforms give rise to many scientific problems…

Multimedia · Computer Science 2025-02-11 Yu Shang , Chen Gao , Nian Li , Yong Li

We introduce Argoverse 2 (AV2) - a collection of three datasets for perception and forecasting research in the self-driving domain. The annotated Sensor Dataset contains 1,000 sequences of multimodal data, encompassing high-resolution…

In recent years, automatic video caption generation has attracted considerable attention. This paper focuses on the generation of Japanese captions for describing human actions. While most currently available video caption datasets have…

Computation and Language · Computer Science 2020-03-11 Yutaro Shigeto , Yuya Yoshikawa , Jiaqing Lin , Akikazu Takeuchi

Autonomous Vehicle (AV) perception systems require more than simply seeing, via e.g., object detection or scene segmentation. They need a holistic understanding of what is happening within the scene for safe interaction with other road…

Computer Vision and Pattern Recognition · Computer Science 2024-11-11 Salman Khan , Izzeddin Teeti , Reza Javanmard Alitappeh , Mihaela C. Stoian , Eleonora Giunchiglia , Gurkirt Singh , Andrew Bradley , Fabio Cuzzolin

Short video platforms, such as YouTube, Instagram, or TikTok, are used by billions of users. These platforms expose users to harmful content, ranging from clickbait or physical harms to hate or misinformation. Yet, we lack a comprehensive…

Computer Vision and Pattern Recognition · Computer Science 2025-04-24 Wonjeong Jo , Magdalena Wojcieszak

Traffic accidents present complex challenges for autonomous driving, often featuring unpredictable scenarios that hinder accurate system interpretation and responses. Nonetheless, prevailing methodologies fall short in elucidating the…

Computer Vision and Pattern Recognition · Computer Science 2025-03-05 Cheng Li , Keyuan Zhou , Tong Liu , Yu Wang , Mingqiao Zhuang , Huan-ang Gao , Bu Jin , Hao Zhao