English
Related papers

Related papers: Automated Dyadic Data Recorder (ADDR) Framework an…

200 papers

We introduce a video framework for modeling the association between verbal and non-verbal communication during dyadic conversation. Given the input speech of a speaker, our approach retrieves a video of a listener, who has facial…

Computer Vision and Pattern Recognition · Computer Science 2023-01-27 Scott Geng , Revant Teotia , Purva Tendulkar , Sachit Menon , Carl Vondrick

We present a framework for modeling interactional communication in dyadic conversations: given multimodal inputs of a speaker, we autoregressively output multiple possibilities of corresponding listener motion. We combine the motion and…

Computer Vision and Pattern Recognition · Computer Science 2022-04-19 Evonne Ng , Hanbyul Joo , Liwen Hu , Hao Li , Trevor Darrell , Angjoo Kanazawa , Shiry Ginosar

Human-human communication is like a delicate dance where listeners and speakers concurrently interact to maintain conversational dynamics. Hence, an effective model for generating listener nonverbal behaviors requires understanding the…

Computer Vision and Pattern Recognition · Computer Science 2024-07-19 Minh Tran , Di Chang , Maksim Siniukov , Mohammad Soleymani

Automated deception detection systems can enhance health, justice, and security in society by helping humans detect deceivers in high-stakes situations across medical and legal domains, among others. This paper presents a novel analysis of…

Computer Vision and Pattern Recognition · Computer Science 2020-10-27 Leena Mathur , Maja J Matarić

Nowadays, automatical personality inference is drawing extensive attention from both academia and industry. Conventional methods are mainly based on user generated contents, e.g., profiles, likes, and texts of an individual, on social…

Computation and Language · Computer Science 2020-09-29 Qiang Liu

Non-contact automatic deception detection remains challenging because visual and auditory deception cues often lack stable cross-subject patterns. In contrast, galvanic skin response (GSR) provides more reliable physiological cues and has…

Computer Vision and Pattern Recognition · Computer Science 2026-04-07 Peiyuan Jiang , Yao Liu , Yanglei Gan , Jiaye Yang , Lu Liu , Daibing Yao , Qiao Liu

The majority of traditional text-to-video retrieval systems operate in static environments, i.e., there is no interaction between the user and the agent beyond the initial textual query provided by the user. This can be sub-optimal if the…

Computer Vision and Pattern Recognition · Computer Science 2022-07-19 Avinash Madasu , Junier Oliva , Gedas Bertasius

In dyadic interactions, humans communicate their intentions and state of mind using verbal and non-verbal cues, where multiple different facial reactions might be appropriate in response to a specific speaker behaviour. Then, how to develop…

Computer Vision and Pattern Recognition · Computer Science 2024-01-11 Siyang Song , Micol Spitale , Cheng Luo , Cristina Palmero , German Barquero , Hengde Zhu , Sergio Escalera , Michel Valstar , Tobias Baur , Fabien Ringeval , Elisabeth Andre , Hatice Gunes

In dyadic interaction, predicting the listener's facial reactions is challenging as different reactions could be appropriate in response to the same speaker's behaviour. Previous approaches predominantly treated this task as an…

Computer Vision and Pattern Recognition · Computer Science 2024-11-05 Cheng Luo , Siyang Song , Weicheng Xie , Micol Spitale , Zongyuan Ge , Linlin Shen , Hatice Gunes

This study investigates the efficacy of using multimodal machine learning techniques to detect deception in dyadic interactions, focusing on the integration of data from both the deceiver and the deceived. We compare early and late fusion…

Machine Learning · Computer Science 2025-12-12 Thomas Jack Samuels , Franco Rugolon , Stephan Hau , Lennart Högman

We present a computational framework for automatically quantifying verbal and nonverbal behaviors in the context of job interviews. The proposed framework is trained by analyzing the videos of 138 interview sessions with 69…

Human-Computer Interaction · Computer Science 2015-04-15 Iftekhar Naim , M. Iftekhar Tanveer , Daniel Gildea , Mohammed , Hoque

To enable more natural face-to-face interactions, conversational agents need to adapt their behavior to their interlocutors. One key aspect of this is generation of appropriate non-verbal behavior for the agent, for example facial gestures,…

Computer Vision and Pattern Recognition · Computer Science 2020-10-26 Patrik Jonell , Taras Kucherenko , Gustav Eje Henter , Jonas Beskow

Natural conversations between humans often involve a large number of non-verbal nuanced expressions, displayed at key times throughout the conversation. Understanding and being able to model these complex interactions is essential for…

Computer Vision and Pattern Recognition · Computer Science 2022-07-12 Renke Wang , Ifeoma Nwogu

Recently, a more challenging state tracking task, Audio-Video Scene-Aware Dialogue (AVSD), is catching an increasing amount of attention among researchers. Different from purely text-based dialogue state tracking, the dialogue in AVSD…

Computation and Language · Computer Science 2020-07-21 Xiangyang Mou , Brandyn Sigouin , Ian Steenstra , Hui Su

Speech-driven facial animation methods usually contain two main classes, 3D and 2D talking face, both of which attract considerable research attention in recent years. However, to the best of our knowledge, the research on 3D talking face…

Computer Vision and Pattern Recognition · Computer Science 2024-04-22 Yixiang Zhuang , Baoping Cheng , Yao Cheng , Yuntao Jin , Renshuai Liu , Chengyang Li , Xuan Cheng , Jing Liao , Juncong Lin

We present the Deception Detection and Physiological Monitoring (DDPM) dataset and initial baseline results on this dataset. Our application context is an interview scenario in which the interviewee attempts to deceive the interviewer on…

Computer Vision and Pattern Recognition · Computer Science 2021-06-15 Jeremy Speth , Nathan Vance , Adam Czajka , Kevin W. Bowyer , Diane Wright , Patrick Flynn

Talking face generation aims to synthesize a face video with precise lip synchronization as well as a smooth transition of facial motion over the entire video via the given speech clip and facial image. Most existing methods mainly focus on…

Computer Vision and Pattern Recognition · Computer Science 2020-05-14 Hao Zhu , Huaibo Huang , Yi Li , Aihua Zheng , Ran He

We present the Conversational Data Retrieval (CDR) benchmark, the first comprehensive test set for evaluating systems that retrieve conversation data for product insights. With 1.6k queries across five analytical tasks and 9.1k…

Computation and Language · Computer Science 2026-02-17 Yohan Lee , Yongwoo Song , Sangyeop Kim

Data scarcity and unreliable self-reporting -- such as concealment or exaggeration -- pose fundamental challenges to psychiatric intake and assessment. We propose a multi-agent synthesis framework that explicitly models patient deception to…

Databases · Computer Science 2026-01-15 Xinyuan Zhang , Zijian Wang , Chang Dao , Juexiao Zhou

Whether an interviewee's honest and deceptive responses can be detected by facial expression signals in videos has been debated and requires further research. We developed deep learning models enabled by computer vision to extract temporal…

Human-Computer Interaction · Computer Science 2026-05-19 Hung-Yue Suen , Kuo-En Hung , Che-Wei Liu , Yu-Sheng Su , Han-Chih Fan
‹ Prev 1 2 3 10 Next ›