English
Related papers

Related papers: Sonification of Facial Actions for Musical Express…

200 papers

Purpose: Surgical scene understanding is key to advancing computer-aided and intelligent surgical systems. Current approaches predominantly rely on visual data or end-to-end learning, which limits fine-grained contextual modeling. This work…

Previous approaches to model and analyze facial expression analysis use three different techniques: facial action units, geometric features and graph based modelling. However, previous approaches have treated these technique separately.…

Computer Vision and Pattern Recognition · Computer Science 2016-06-03 Mehdi Ghayoumi , Arvind K Bansal

Automatic recognition of facial gestures is becoming increasingly important as real world AI agents become a reality. In this paper, we present an automated system that recognizes facial gestures by capturing local changes and encoding the…

Computer Vision and Pattern Recognition · Computer Science 2017-06-13 Nadav Israel , Lior Wolf , Ran Barzilay , Gal Shoval

The hand gestures are one of the typical methods used in sign language. It is very difficult for the hearing-impaired people to communicate with the world. This project presents a solution that will not only automatically recognize the hand…

Computer Vision and Pattern Recognition · Computer Science 2018-11-30 K. Manikandan , Ayush Patidar , Pallav Walia , Aneek Barman Roy

The facial expression recognition is an ocular task that can be performed without human discomfort, is really a speedily growing on the computer research field. There are many applications and programs uses facial expression to evaluate…

Computer Vision and Pattern Recognition · Computer Science 2015-06-08 Abdelmajid Hassan Mansour , Gafar Zen Alabdeen Salh , Ali Shaif Alhalemi

Facial expressions are an important way through which humans interact socially. Building a system capable of automatically recognizing facial expressions from images and video has been an intense field of study in recent years. Interpreting…

Computer Vision and Pattern Recognition · Computer Science 2016-06-13 Ciprian Corneanu , Marc Oliu , Jeffrey F. Cohn , Sergio Escalera

As text-to-speech technologies achieve remarkable naturalness in read-aloud tasks, there is growing interest in multimodal synthesis of verbal and non-verbal communicative behaviour, such as spontaneous speech and associated body gestures.…

Audio and Speech Processing · Electrical Eng. & Systems 2024-01-11 Shivam Mehta , Ruibo Tu , Simon Alexanderson , Jonas Beskow , Éva Székely , Gustav Eje Henter

Affective computing plays a key role in human-computer interactions, entertainment, teaching, safe driving, and multimedia integration. Major breakthroughs have been made recently in the areas of affective computing (i.e., emotion…

Multimedia · Computer Science 2022-03-22 Yan Wang , Wei Song , Wei Tao , Antonio Liotta , Dawei Yang , Xinlei Li , Shuyong Gao , Yixuan Sun , Weifeng Ge , Wei Zhang , Wenqiang Zhang

Current audio-driven facial animation methods achieve impressive results for short videos but suffer from error accumulation and identity drift when extended to longer durations. Existing methods attempt to mitigate this through external…

Humans possess a remarkable ability to integrate auditory and visual information, enabling a deeper understanding of the surrounding environment. This early fusion of audio and visual cues, demonstrated through cognitive psychology and…

Computer Vision and Pattern Recognition · Computer Science 2023-12-05 Shentong Mo , Pedro Morgado

Face synthesis is an important problem in computer vision with many applications. In this work, we describe a new method, namely LandmarkGAN, to synthesize faces based on facial landmarks as input. Facial landmarks are a natural, intuitive,…

Computer Vision and Pattern Recognition · Computer Science 2021-02-09 Pu Sun , Yuezun Li , Honggang Qi , Siwei Lyu

Speech-driven facial animation aims to synthesize lip-synchronized 3D talking faces following the given speech signal. Prior methods to this task mostly focus on pursuing realism with deterministic systems, yet characterizing the…

Computer Vision and Pattern Recognition · Computer Science 2024-10-01 Chunzhi Gu , Shigeru Kuriyama , Katsuya Hotta

Audiovisual active speaker detection (ASD) is conventionally performed by modelling the temporal synchronisation of acoustic and visual speech cues. In egocentric recordings, however, the efficacy of synchronisation-based methods is…

Multimedia · Computer Science 2025-06-24 Jason Clarke , Yoshihiko Gotoh , Stefan Goetze

Facial expressions are one of the most powerful, natural and immediate means for human being to communicate their emotions and intensions. Recognition of facial expression has many applications including human-computer interaction,…

Computer Vision and Pattern Recognition · Computer Science 2016-04-18 Deepak Ghimire , Sunghwan Jeong , Joonwhoan Lee , Sang Hyun Park

This paper explores the integration of visual communication and musical interaction by implementing a robotic camera within a "Guided Harmony" musical game. We aim to examine co-creative behaviors between human musicians and robotic…

Human-Computer Interaction · Computer Science 2024-10-29 Ross Greer , Laura Fleig , Shlomo Dubnov

Audio-driven emotional 3D face animation aims to generate emotionally expressive talking heads with synchronized lip movements. However, previous research has often overlooked the influence of diverse emotions on facial expressions or…

Computer Vision and Pattern Recognition · Computer Science 2024-08-06 Chang Liu , Qunfen Lin , Zijiao Zeng , Ye Pan

This chapter presents an overview of Interactive Machine Learning (IML) techniques applied to the analysis and design of musical gestures. We go through the main challenges and needs related to capturing, analysing, and applying IML…

Machine Learning · Computer Science 2020-11-30 Federico Ghelli Visi , Atau Tanaka

This article describes the development of a platform designed to visualize the 3D motion of the tongue using ultrasound image sequences. An overview of the system design is given and promising results are presented. Compared to the analysis…

Computer Vision and Pattern Recognition · Computer Science 2016-05-20 Kele Xu , Yin Yang , Aurore Jaumard-Hakoun , Clemence Leboullenger , Gerard Dreyfus , Pierre Roussel , Maureen Stone , Bruce Denby

We present a framework for generating full-bodied photorealistic avatars that gesture according to the conversational dynamics of a dyadic interaction. Given speech audio, we output multiple possibilities of gestural motion for an…

Computer Vision and Pattern Recognition · Computer Science 2024-01-04 Evonne Ng , Javier Romero , Timur Bagautdinov , Shaojie Bai , Trevor Darrell , Angjoo Kanazawa , Alexander Richard

The ability to modulate vocal sounds and generate speech is one of the features which set humans apart from other living beings. The human voice can be characterized by several attributes such as pitch, timbre, loudness, and vocal tone. It…

Computer Vision and Pattern Recognition · Computer Science 2017-10-30 Poorna Banerjee Dasgupta