English
Related papers

Related papers: YouTube AV 50K: An Annotated Corpus for Comments i…

200 papers

As the popularity of autonomous vehicles has grown, many standards and regulators, such as ISO, NHTSA, and Euro NCAP, require safety validation to ensure a sufficient level of safety before deploying them in the real world. Manufacturers…

Machine Learning · Computer Science 2024-10-30 Linh Trinh , Ali Anwar , Siegfried Mercelis

With the rapid advancement of video generation models such as Sora, video quality assessment (VQA) is becoming increasingly crucial for selecting high-quality videos from large-scale datasets used in pre-training. Traditional VQA methods,…

Computer Vision and Pattern Recognition · Computer Science 2025-09-16 Yanyun Pu , Kehan Li , Zeyi Huang , Zhijie Zhong , Kaixiang Yang

Content creators increasingly utilize generative artificial intelligence (Gen-AI) on platforms such as YouTube, TikTok, Instagram, and various blogging sites to produce imaginative images, AI-generated videos, and articles using Large…

Human-Computer Interaction · Computer Science 2024-03-12 Yao Lyu , He Zhang , Shuo Niu , Jie Cai

Millions of people use platforms such as YouTube, Facebook, Twitter, and other mass media. Due to the accessibility of these platforms, they are often used to establish a narrative, conduct propaganda, and disseminate misinformation. This…

Machine Learning · Computer Science 2021-07-05 Raj Jagtap , Abhinav Kumar , Rahul Goel , Shakshi Sharma , Rajesh Sharma , Clint P. George

Visual Question Answering (VQA) entails answering questions about images. We introduce the first VQA dataset in which all contents originate from an authentic use case. Sourced from online question answering community forums, we call it…

Computer Vision and Pattern Recognition · Computer Science 2024-07-18 Chongyan Chen , Mengchen Liu , Noel Codella , Yunsheng Li , Lu Yuan , Danna Gurari

Publishing open-source academic video recordings is an emergent and prevalent approach to sharing knowledge online. Such videos carry rich multimodal information including speech, the facial and body movements of the speakers, as well as…

Computation and Language · Computer Science 2024-06-05 Zhe Chen , Heyang Liu , Wenyi Yu , Guangzhi Sun , Hongcheng Liu , Ji Wu , Chao Zhang , Yu Wang , Yanfeng Wang

Video communication has been rapidly increasing over the past decade, with YouTube providing a medium where users can post, discover, share, and react to videos. There has also been an increase in the number of videos citing research…

Digital Libraries · Computer Science 2022-09-07 Abdul Rahman Shaikh , Hamed Alhoori , Maoyuan Sun

As short-form funny videos on social networks are gaining popularity, it becomes demanding for AI models to understand them for better communication with humans. Unfortunately, previous video humor datasets target specific domains, such as…

Computation and Language · Computer Science 2024-04-02 Dayoon Ko , Sangho Lee , Gunhee Kim

This paper describes the AVA-Kinetics localized human actions video dataset. The dataset is collected by annotating videos from the Kinetics-700 dataset using the AVA annotation protocol, and extending the original AVA dataset with these…

Computer Vision and Pattern Recognition · Computer Science 2020-05-21 Ang Li , Meghana Thotakuri , David A. Ross , João Carreira , Alexander Vostrikov , Andrew Zisserman

The objective of this study is to investigate automated vehicle (AV) adoption perceptions, including ownership intentions and the willingness to use self-driving mobility services. In this paper, we use data from the 2018 California…

Computers and Society · Computer Science 2024-07-18 Tho Le , Giovanni Circella

Video content is rich in semantics and has the ability to evoke various emotions in viewers. In recent years, with the rapid development of affective computing and the explosive growth of visual data, affective video content analysis (AVCA)…

Computer Vision and Pattern Recognition · Computer Science 2024-01-19 Junxiao Xue , Jie Wang , Xuecheng Wu , Qian Zhang

We present ClimateChat-300K, a large-scale dataset of 299,329 public Facebook posts about climate change collected between May 2020 and May 2024 through the CrowdTangle platform. The dataset contains 41 metadata features including post…

Computation and Language · Computer Science 2026-05-25 Wajdi Zaghouani , Md. Rafiul Biswas , Mabrouka Bessghaier , Shimaa Ibrahim , George Mikros

In recent years, vision-centric perception has flourished in various autonomous driving tasks, including 3D detection, semantic map construction, motion forecasting, and depth estimation. Nevertheless, the latency of vision-centric…

Computer Vision and Pattern Recognition · Computer Science 2022-12-20 Xiaofeng Wang , Zheng Zhu , Yunpeng Zhang , Guan Huang , Yun Ye , Wenbo Xu , Ziwei Chen , Xingang Wang

Automatic sentiment analysis play vital role in decision making. Many organizations spend a lot of budget to understand their customer satisfaction by manually going over their feedback/comments or tweets. Automatic sentiment analysis can…

Computation and Language · Computer Science 2021-07-07 Mohammad Aimal , Maheen Bakhtyar , Junaid Baber , Sadia Lakho , Umar Mohammad , Warda Ahmed , Jahanvash Karim

The availability of high definition video content on the web has brought about a significant change in the characteristics of Internet video, but not many studies on characterizing video have been done after this change. Video…

Multimedia · Computer Science 2014-08-26 Saba Ahsan , Varun Singh , Jörg Ott

Large vision-language models (VLMs) have garnered increasing interest in autonomous driving areas, due to their advanced capabilities in complex reasoning tasks essential for highly autonomous vehicle behavior. Despite their potential,…

Computer Vision and Pattern Recognition · Computer Science 2024-07-23 Ming Nie , Renyuan Peng , Chunwei Wang , Xinyue Cai , Jianhua Han , Hang Xu , Li Zhang

We present a new dataset with annotated eye movements. The dataset consists of over 800,000 gaze points recorded during a car ride in the real world and in the simulator. In total, the eye movements of 19 subjects were annotated. In this…

Computer Vision and Pattern Recognition · Computer Science 2021-01-13 Wolfgang Fuhl , Enkelejda Kasneci

Growing interest in autonomous driving (AD) and intelligent vehicles (IVs) is fueled by their promise for enhanced safety, efficiency, and economic benefits. While previous surveys have captured progress in this field, a comprehensive and…

Robotics · Computer Science 2023-06-12 Long Chen , Siyu Teng , Bai Li , Xiaoxiang Na , Yuchen Li , Zixuan Li , Jinjun Wang , Dongpu Cao , Nanning Zheng , Fei-Yue Wang

In recent years, deep neural networks have demonstrated increasingly strong abilities to recognize objects and activities in videos. However, as video understanding becomes widely used in real-world applications, a key consideration is…

Computer Vision and Pattern Recognition · Computer Science 2022-10-19 Mantas Mazeika , Eric Tang , Andy Zou , Steven Basart , Jun Shern Chan , Dawn Song , David Forsyth , Jacob Steinhardt , Dan Hendrycks

Satellites are capable of capturing high-resolution videos. It makes vehicle perception from satellite become possible. Compared to street surveillance, drive recorder or other equipments, satellite videos provide a much broader city-scale…

Computer Vision and Pattern Recognition · Computer Science 2024-03-27 Bin Zhao , Pengfei Han , Xuelong Li