中文
相关论文

相关论文: YouTube AV 50K: An Annotated Corpus for Comments i…

200 篇论文

As the popularity of autonomous vehicles has grown, many standards and regulators, such as ISO, NHTSA, and Euro NCAP, require safety validation to ensure a sufficient level of safety before deploying them in the real world. Manufacturers…

机器学习 · 计算机科学 2024-10-30 Linh Trinh , Ali Anwar , Siegfried Mercelis

With the rapid advancement of video generation models such as Sora, video quality assessment (VQA) is becoming increasingly crucial for selecting high-quality videos from large-scale datasets used in pre-training. Traditional VQA methods,…

计算机视觉与模式识别 · 计算机科学 2025-09-16 Yanyun Pu , Kehan Li , Zeyi Huang , Zhijie Zhong , Kaixiang Yang

Content creators increasingly utilize generative artificial intelligence (Gen-AI) on platforms such as YouTube, TikTok, Instagram, and various blogging sites to produce imaginative images, AI-generated videos, and articles using Large…

人机交互 · 计算机科学 2024-03-12 Yao Lyu , He Zhang , Shuo Niu , Jie Cai

Millions of people use platforms such as YouTube, Facebook, Twitter, and other mass media. Due to the accessibility of these platforms, they are often used to establish a narrative, conduct propaganda, and disseminate misinformation. This…

机器学习 · 计算机科学 2021-07-05 Raj Jagtap , Abhinav Kumar , Rahul Goel , Shakshi Sharma , Rajesh Sharma , Clint P. George

Visual Question Answering (VQA) entails answering questions about images. We introduce the first VQA dataset in which all contents originate from an authentic use case. Sourced from online question answering community forums, we call it…

计算机视觉与模式识别 · 计算机科学 2024-07-18 Chongyan Chen , Mengchen Liu , Noel Codella , Yunsheng Li , Lu Yuan , Danna Gurari

Publishing open-source academic video recordings is an emergent and prevalent approach to sharing knowledge online. Such videos carry rich multimodal information including speech, the facial and body movements of the speakers, as well as…

计算与语言 · 计算机科学 2024-06-05 Zhe Chen , Heyang Liu , Wenyi Yu , Guangzhi Sun , Hongcheng Liu , Ji Wu , Chao Zhang , Yu Wang , Yanfeng Wang

Video communication has been rapidly increasing over the past decade, with YouTube providing a medium where users can post, discover, share, and react to videos. There has also been an increase in the number of videos citing research…

数字图书馆 · 计算机科学 2022-09-07 Abdul Rahman Shaikh , Hamed Alhoori , Maoyuan Sun

As short-form funny videos on social networks are gaining popularity, it becomes demanding for AI models to understand them for better communication with humans. Unfortunately, previous video humor datasets target specific domains, such as…

计算与语言 · 计算机科学 2024-04-02 Dayoon Ko , Sangho Lee , Gunhee Kim

This paper describes the AVA-Kinetics localized human actions video dataset. The dataset is collected by annotating videos from the Kinetics-700 dataset using the AVA annotation protocol, and extending the original AVA dataset with these…

计算机视觉与模式识别 · 计算机科学 2020-05-21 Ang Li , Meghana Thotakuri , David A. Ross , João Carreira , Alexander Vostrikov , Andrew Zisserman

The objective of this study is to investigate automated vehicle (AV) adoption perceptions, including ownership intentions and the willingness to use self-driving mobility services. In this paper, we use data from the 2018 California…

计算机与社会 · 计算机科学 2024-07-18 Tho Le , Giovanni Circella

Video content is rich in semantics and has the ability to evoke various emotions in viewers. In recent years, with the rapid development of affective computing and the explosive growth of visual data, affective video content analysis (AVCA)…

计算机视觉与模式识别 · 计算机科学 2024-01-19 Junxiao Xue , Jie Wang , Xuecheng Wu , Qian Zhang

We present ClimateChat-300K, a large-scale dataset of 299,329 public Facebook posts about climate change collected between May 2020 and May 2024 through the CrowdTangle platform. The dataset contains 41 metadata features including post…

计算与语言 · 计算机科学 2026-05-25 Wajdi Zaghouani , Md. Rafiul Biswas , Mabrouka Bessghaier , Shimaa Ibrahim , George Mikros

In recent years, vision-centric perception has flourished in various autonomous driving tasks, including 3D detection, semantic map construction, motion forecasting, and depth estimation. Nevertheless, the latency of vision-centric…

计算机视觉与模式识别 · 计算机科学 2022-12-20 Xiaofeng Wang , Zheng Zhu , Yunpeng Zhang , Guan Huang , Yun Ye , Wenbo Xu , Ziwei Chen , Xingang Wang

Automatic sentiment analysis play vital role in decision making. Many organizations spend a lot of budget to understand their customer satisfaction by manually going over their feedback/comments or tweets. Automatic sentiment analysis can…

计算与语言 · 计算机科学 2021-07-07 Mohammad Aimal , Maheen Bakhtyar , Junaid Baber , Sadia Lakho , Umar Mohammad , Warda Ahmed , Jahanvash Karim

The availability of high definition video content on the web has brought about a significant change in the characteristics of Internet video, but not many studies on characterizing video have been done after this change. Video…

多媒体 · 计算机科学 2014-08-26 Saba Ahsan , Varun Singh , Jörg Ott

Large vision-language models (VLMs) have garnered increasing interest in autonomous driving areas, due to their advanced capabilities in complex reasoning tasks essential for highly autonomous vehicle behavior. Despite their potential,…

计算机视觉与模式识别 · 计算机科学 2024-07-23 Ming Nie , Renyuan Peng , Chunwei Wang , Xinyue Cai , Jianhua Han , Hang Xu , Li Zhang

We present a new dataset with annotated eye movements. The dataset consists of over 800,000 gaze points recorded during a car ride in the real world and in the simulator. In total, the eye movements of 19 subjects were annotated. In this…

计算机视觉与模式识别 · 计算机科学 2021-01-13 Wolfgang Fuhl , Enkelejda Kasneci

Growing interest in autonomous driving (AD) and intelligent vehicles (IVs) is fueled by their promise for enhanced safety, efficiency, and economic benefits. While previous surveys have captured progress in this field, a comprehensive and…

机器人学 · 计算机科学 2023-06-12 Long Chen , Siyu Teng , Bai Li , Xiaoxiang Na , Yuchen Li , Zixuan Li , Jinjun Wang , Dongpu Cao , Nanning Zheng , Fei-Yue Wang

In recent years, deep neural networks have demonstrated increasingly strong abilities to recognize objects and activities in videos. However, as video understanding becomes widely used in real-world applications, a key consideration is…

计算机视觉与模式识别 · 计算机科学 2022-10-19 Mantas Mazeika , Eric Tang , Andy Zou , Steven Basart , Jun Shern Chan , Dawn Song , David Forsyth , Jacob Steinhardt , Dan Hendrycks

Satellites are capable of capturing high-resolution videos. It makes vehicle perception from satellite become possible. Compared to street surveillance, drive recorder or other equipments, satellite videos provide a much broader city-scale…

计算机视觉与模式识别 · 计算机科学 2024-03-27 Bin Zhao , Pengfei Han , Xuelong Li