English
Related papers

Related papers: Cross-artform performance using networked interfac…

200 papers

Human action or activity recognition in videos is a fundamental task in computer vision with applications in surveillance and monitoring, self-driving cars, sports analytics, human-robot interaction and many more. Traditional supervised…

Computer Vision and Pattern Recognition · Computer Science 2024-04-10 Sharana Dharshikgan Suresh Dass , Hrishav Bakul Barua , Ganesh Krishnasamy , Raveendran Paramesran , Raphael C. -W. Phan

Image-based virtual try-on aims to transfer an in-shop clothing image to a person image. Most existing methods adopt a single global deformation to perform clothing warping directly, which lacks fine-grained modeling of in-shop clothing and…

Computer Vision and Pattern Recognition · Computer Science 2025-03-19 Shengping Zhang , Xiaoyu Han , Weigang Zhang , Xiangyuan Lan , Hongxun Yao , Qingming Huang

Emotion represents an essential aspect of human speech that is manifested in speech prosody. Speech, visual, and textual cues are complementary in human communication. In this paper, we study a hybrid fusion method, referred to as…

Audio and Speech Processing · Electrical Eng. & Systems 2020-09-10 Zexu Pan , Zhaojie Luo , Jichen Yang , Haizhou Li

Current state-of-the-art methods solve spatiotemporal action localisation by extending 2D anchors to 3D-cuboid proposals on stacks of frames, to generate sets of temporally connected bounding boxes called \textit{action micro-tubes}.…

Image and Video Processing · Electrical Eng. & Systems 2018-08-02 Gurkirt Singh , Suman Saha , Fabio Cuzzolin

This paper uses the capabilities of latent diffusion models (LDMs) to generate realistic RGB human-object interaction scenes to guide humanoid loco-manipulation planning. To do so, we extract from the generated images both the contact…

Robotics · Computer Science 2025-04-24 Ilyass Taouil , Haizhou Zhao , Angela Dai , Majid Khadiv

We address the challenging task of human reaction generation, which aims to generate a corresponding reaction based on an input action. Most of the existing works do not focus on generating and predicting the reaction and cannot generate…

Computer Vision and Pattern Recognition · Computer Science 2023-02-03 Baptiste Chopin , Hao Tang , Naima Otberdout , Mohamed Daoudi , Nicu Sebe

Due to the inter- and intra- variation of respiratory motion, it is highly desired to provide real-time volumetric images during the treatment delivery of lung stereotactic body radiation therapy (SBRT) for accurate and active motion…

We present the first experimental demonstration of a neuromorphic network with magnetic tunnel junction (MTJ) synapses, which performs image recognition via vector-matrix multiplication. We also simulate a large MTJ network performing MNIST…

Neural and Evolutionary Computing · Computer Science 2021-12-15 Peng Zhou , Alexander J. Edwards , Fred B. Mancoff , Dimitri Houssameddine , Sanjeev Aggarwal , Joseph S. Friedman

Accurate modeling of moving boundaries and interfaces is a difficulty present in many situations of computational mechanics. We use the eXtreme Mesh deformation approach (X-Mesh) to simulate the interaction between two immiscible flows…

Computational Engineering, Finance, and Science · Computer Science 2024-02-02 Antoine Quiriny , Jonathan Lambrechts , Nicolas Moës , Jean-François Remacle

Vision Transformers (ViTs) have shown strong empirical performance on high-dimensional medical imaging data, yet their behavior under survival objectives and the interpretability of their attention mechanisms remain poorly understood. Under…

Medical Physics · Physics 2026-04-24 Qiyuan Shi , Yi Li

In this paper, we propose a novel representation for grasping using contacts between multi-finger robotic hands and objects to be manipulated. This representation significantly reduces the prediction dimensions and accelerates the learning…

This paper introduces ActGAN - a novel end-to-end generative adversarial network (GAN) for one-shot face reenactment. Given two images, the goal is to transfer the facial expression of the source actor onto a target person in a…

Computer Vision and Pattern Recognition · Computer Science 2020-04-01 Ivan Kosarevych , Marian Petruk , Markian Kostiv , Orest Kupyn , Mykola Maksymenko , Volodymyr Budzan

The paper presents a pilot exploration of the construction, management and analysis of a multimodal corpus. Through a three-layer annotation that provides orthographic, prosodic, and gestural transcriptions, the Gest-IT resource allows to…

Recognizing human activities in videos is challenging due to the spatio-temporal complexity and context-dependence of human interactions. Prior studies often rely on single input modalities, such as RGB or skeletal data, limiting their…

Computer Vision and Pattern Recognition · Computer Science 2024-09-05 Tuyen Tran , Thao Minh Le , Hung Tran , Truyen Tran

In many contexts, creating mappings for gestural interactions can form part of an artistic process. Creators seeking a mapping that is expressive, novel, and affords them a sense of authorship may not know how to program it up in a signal…

Human-Computer Interaction · Computer Science 2021-06-17 Tim Murray-Browne , Panagiotis Tigas

The proposed humanistic approach mapped the human character and behavior into a device to evade the bondages of implementation and surely succeed as we live. Human societies are the complex and most organized networks, in which many…

Networking and Internet Architecture · Computer Science 2013-11-14 Md. Amir Khusru Akhtar , G. Sahoo

This paper presents I-nteract, a cyber-physical system that enables real-time interaction with real and virtual objects in a mixed augmented reality environment to design 3D models for additive manufacturing. The system has been developed…

Human-Computer Interaction · Computer Science 2020-02-18 Ammar Malik , Hugo Lhachemi , Robert Shorten

User Interaction for NFTs (Non-fungible Tokens) is gaining increasing attention. Although NFTs have been traditionally single-use and monolithic, recent applications aim to connect multimodal interaction with human behaviour. This paper…

Multimedia · Computer Science 2022-06-09 Anqi Wang , Ze Gao , Lik-Hang Lee , Tristan Braud , Pan Hui

Actions are about how we interact with the environment, including other people, objects, and ourselves. In this paper, we propose a novel multi-modal Holistic Interaction Transformer Network (HIT) that leverages the largely ignored, but…

Computer Vision and Pattern Recognition · Computer Science 2022-11-21 Gueter Josmy Faure , Min-Hung Chen , Shang-Hong Lai

We introduce X-Streamer, an end-to-end multimodal human world modeling framework for building digital human agents capable of infinite interactions across text, speech, and video within a single unified architecture. Starting from a single…

Computer Vision and Pattern Recognition · Computer Science 2025-09-29 You Xie , Tianpei Gu , Zenan Li , Chenxu Zhang , Guoxian Song , Xiaochen Zhao , Chao Liang , Jianwen Jiang , Hongyi Xu , Linjie Luo