中文
相关论文

相关论文: VEATIC: Video-based Emotion and Affect Tracking in…

200 篇论文

Recent advances in AI-generated video have shown strong performance on \emph{text-to-video} tasks, particularly for short clips depicting a single scene. However, current models struggle to generate longer videos with coherent scene…

计算机视觉与模式识别 · 计算机科学 2026-02-03 Hanwen Shen , Jiajie Lu , Yupeng Cao , Xiaonan Yang

The primary objective is to teach a machine about human emotions, which has become an essential requirement in the field of social intelligence, also expedites the progress of human-machine interactions. The ability of a machine to…

音频与语音处理 · 电气工程与系统科学 2020-06-23 Sai Nikhil Chennoor , B. R. K. Madhur , Moujiz Ali , T. Kishore Kumar

Automatic understanding of human affect using visual signals is a problem that has attracted significant interest over the past 20 years. However, human emotional states are quite complex. To appraise such states displayed in real-world…

计算机视觉与模式识别 · 计算机科学 2019-12-17 Dimitrios Kollias , Stefanos Zafeiriou

We present \textbf{FakeET}-- an eye-tracking database to understand human visual perception of \emph{deepfake} videos. Given that the principal purpose of deepfakes is to deceive human observers, FakeET is designed to understand and…

计算机视觉与模式识别 · 计算机科学 2020-06-22 Parul Gupta , Komal Chugh , Abhinav Dhall , Ramanathan Subramanian

The scarcity of high-quality large-scale labeled datasets poses a huge challenge for employing deep learning models in video deception detection. To address this issue, inspired by the psychological theory on the relation between deception…

计算机视觉与模式识别 · 计算机科学 2024-12-13 Zihan Ji , Xuetao Tian , Ye Liu

Video-based human action recognition is currently one of the most active research areas in computer vision. Various research studies indicate that the performance of action recognition is highly dependent on the type of features being…

计算机视觉与模式识别 · 计算机科学 2019-10-02 Lei Wang , Du Q. Huynh , Piotr Koniusz

Image Captioning, the task of automatic generation of image captions, has attracted attentions from researchers in many fields of computer science, being computer vision, natural language processing and machine learning in recent years.…

计算与语言 · 计算机科学 2020-02-04 Quan Hoang Lam , Quang Duy Le , Kiet Van Nguyen , Ngan Luu-Thuy Nguyen

Emotional content is a crucial ingredient in user-generated videos. However, the sparsity of emotional expressions in the videos poses an obstacle to visual emotion analysis. In this paper, we propose a new neural approach, Bi-stream…

机器学习 · 计算机科学 2019-07-25 Guoyun Tu , Yanwei Fu , Boyang Li , Jiarui Gao , Yu-Gang Jiang , Xiangyang Xue

We present Affect2MM, a learning method for time-series emotion prediction for multimedia content. Our goal is to automatically capture the varying emotions depicted by characters in real-life human-centric situations and behaviors. We use…

计算机视觉与模式识别 · 计算机科学 2021-03-12 Trisha Mittal , Puneet Mathur , Aniket Bera , Dinesh Manocha

Text-level discourse parsing aims to unmask how two sentences in the text are related to each other. We propose the task of Visual Discourse Parsing, which requires understanding discourse relations among scenes in a video. Here we use the…

计算机视觉与模式识别 · 计算机科学 2022-01-25 Arjun R. Akula , Song-Chun Zhu

Real-time visual feedback from catheterization analysis is crucial for enhancing surgical safety and efficiency during endovascular interventions. However, existing datasets are often limited to specific tasks, small scale, and lack the…

Humans are arguably one of the most important subjects in video streams, many real-world applications such as video summarization or video editing workflows often require the automatic search and retrieval of a person of interest. Despite…

计算机视觉与模式识别 · 计算机科学 2021-06-04 Juan Leon Alcazar , Long Mai , Federico Perazzi , Joon-Young Lee , Pablo Arbelaez , Bernard Ghanem , Fabian Caba Heilbron

Automatic affect recognition using visual cues is an important task towards a complete interaction between humans and machines. Applications can be found in tutoring systems and human computer interaction. A critical step towards that…

计算机视觉与模式识别 · 计算机科学 2021-11-05 Panagiotis Tzirakis , Dénes Boros , Elnar Hajiyev , Björn W. Schuller

In crowd behavior understanding, a model of crowd behavior need to be trained using the information extracted from video sequences. Since there is no ground-truth available in crowd datasets except the crowd behavior labels, most of the…

计算机视觉与模式识别 · 计算机科学 2016-07-27 Hamidreza Rabiee , Javad Haddadnia , Hossein Mousavi , Moin Nabi , Vittorio Murino , Nicu Sebe

Existing affective-computing, social-signal-processing, and meeting corpora capture important parts of human interaction, but they rarely support analysis of affect in co-located groups as a coupled individual, interpersonal, and…

Early detection of psychological distress is key to effective treatment. Automatic detection of distress, such as depression, is an active area of research. Current approaches utilise vocal, facial, and bodily modalities. Of these, the…

计算机视觉与模式识别 · 计算机科学 2020-03-03 Indigo J. D. Orton

The continuous and increasing use of social media has enabled the expression of human thoughts, opinions, and everyday actions publicly at an unprecedented scale. We present the Vent dataset, the largest annotated dataset of text, emotions,…

社会与信息网络 · 计算机科学 2019-03-26 Nikolaos Lykousas , Costantinos Patsakis , Andreas Kaltenbrunner , Vicenç Gómez

Human communication involves a complex interplay of verbal and nonverbal signals, essential for conveying meaning and achieving interpersonal goals. To develop socially intelligent AI technologies, it is crucial to develop models that can…

Currently, studying the vehicle-human interactive behavior in the emergency needs a large amount of datasets in the actual emergent situations that are almost unavailable. Existing public data sources on autonomous vehicles (AVs) mainly…

计算机视觉与模式识别 · 计算机科学 2020-08-13 Wansong Liu , Danyang Luo , Changxu Wu , Minghui Zheng

The volumetric representation of human interactions is one of the fundamental domains in the development of immersive media productions and telecommunication applications. Particularly in the context of the rapid advancement of Extended…

计算机视觉与模式识别 · 计算机科学 2024-02-15 Fatemeh Ghorbani Lohesara , Davi Rabbouni Freitas , Christine Guillemot , Karen Eguiazarian , Sebastian Knorr