English
Related papers

Related papers: QUB-PHEO: A Visual-Based Dyadic Multi-View Dataset…

200 papers

We propose a new task to benchmark human-in-scene understanding for embodied agents: Human-In-Scene Question Answering (HIS-QA). Given a human motion within a 3D scene, HIS-QA requires the agent to comprehend human states and behaviors,…

Computer Vision and Pattern Recognition · Computer Science 2025-07-17 Jiahe Zhao , Ruibing Hou , Zejie Tian , Hong Chang , Shiguang Shan

This paper presents a new large multiview dataset called HUMBI for human body expressions with natural clothing. The goal of HUMBI is to facilitate modeling view-specific appearance and geometry of five primary body signals including gaze,…

Computer Vision and Pattern Recognition · Computer Science 2021-12-22 Jae Shin Yoon , Zhixuan Yu , Jaesik Park , Hyun Soo Park

We present a motion planning algorithm to compute collision-free and smooth trajectories for high-DOF robots interacting with humans in a shared workspace. Our approach uses offline learning of human actions along with temporal coherence to…

Robotics · Computer Science 2017-11-28 Jae Sung Park , Chonhyon Park , Dinesh Manocha

Providing artificial agents with the same computational models of biological systems is a way to understand how intelligent behaviours may emerge. We present an active inference body perception and action model working for the first time in…

Robotics · Computer Science 2021-02-08 Guillermo Oliver , Pablo Lanillos , Gordon Cheng

We introduce BusyBoard, a toy-inspired robot learning environment that leverages a diverse set of articulated objects and inter-object functional relations to provide rich visual feedback for robot interactions. Based on this environment,…

Robotics · Computer Science 2022-10-18 Zeyi Liu , Zhenjia Xu , Shuran Song

The effectiveness of human-robot interaction often hinges on the ability to cultivate engagement - a dynamic process of cognitive involvement that supports meaningful exchanges. Many existing definitions and models of engagement are either…

Robotics · Computer Science 2025-12-04 Dominykas Strazdas , Magnus Jung , Jan Marquenie , Ingo Siegert , Ayoub Al-Hamadi

Rapid progress has been witnessed for human-object interaction (HOI) recognition, but most existing models are confined to single-stage reasoning pipelines. Considering the intrinsic complexity of the task, we introduce a cascade…

Computer Vision and Pattern Recognition · Computer Science 2020-03-26 Tianfei Zhou , Wenguan Wang , Siyuan Qi , Haibin Ling , Jianbing Shen

Livestreaming often involves interactions between streamers and objects, which is critical for understanding and regulating web content. While human-object interaction (HOI) detection has made some progress in general-purpose video…

Computer Vision and Pattern Recognition · Computer Science 2025-08-05 Menghui Zhang , Jing Zhang , Lin Chen , Li Zhuo

We present HRDexDB, a large-scale, multi-modal dataset of high-fidelity dexterous grasping sequences featuring both human and diverse robotic hands. Unlike existing datasets, HRDexDB provides a comprehensive collection of grasping…

Robotics · Computer Science 2026-04-17 Jongbin Lim , Taeyun Ha , Mingi Choi , Jisoo Kim , Byungjun Kim , Subin Jeon , Hanbyul Joo

Humanoid robots have shown success in locomotion and manipulation. Despite these basic abilities, humanoids are still required to quickly understand human instructions and react based on human interaction signals to become valuable…

Learning physical interaction skills, such as dancing, handshaking, or sparring, remains a fundamental challenge for agents operating in human environments, particularly when the agent's morphology differs significantly from that of the…

Robotics · Computer Science 2025-08-05 Tianyu Li , Hengbo Ma , Sehoon Ha , Kwonjoon Lee

Modeling interaction dynamics to generate robot trajectories that enable a robot to adapt and react to a human's actions and intentions is critical for efficient and effective collaborative Human-Robot Interactions (HRI). Learning from…

Robotics · Computer Science 2023-01-24 Vignesh Prasad , Dorothea Koert , Ruth Stock-Homburg , Jan Peters , Georgia Chalvatzaki

The growing presence of service robots in human-centric environments, such as warehouses, demands seamless and intuitive human-robot collaboration. In this paper, we propose a collaborative shelf-picking framework that combines multimodal…

Robotics · Computer Science 2025-04-10 Abhinav Pathak , Kalaichelvi Venkatesan , Tarek Taha , Rajkumar Muthusamy

This work presents novel robot-mediated immersive experiences enabled by an encountered-type haptic display (ETHD) that introduces direct physical contact in virtual environments. We focus on social-physical interactions, a class of…

Human-Computer Interaction · Computer Science 2026-02-13 Eric Godden , Jacquie Groenewegen , Michael Wheeler , Matthew K. X. J. Pan

Robotic task planning in real-world environments requires not only object recognition but also a nuanced understanding of spatial relationships between objects. We present a spatial-relationship-aware dataset of nearly 1,000 robot-acquired…

Robotics · Computer Science 2025-06-17 Peng Wang , Minh Huy Pham , Zhihao Guo , Wei Zhou

Human-object interactions (HOI) recognition and pose estimation are two closely related tasks. Human pose is an essential cue for recognizing actions and localizing the interacted objects. Meanwhile, human action and their interacted…

Computer Vision and Pattern Recognition · Computer Science 2019-03-18 Wei Feng , Wentao Liu , Tong Li , Jing Peng , Chen Qian , Xiaolin Hu

Current vision-language multimodal models are well-adapted for general visual understanding tasks. However, they perform inadequately when handling complex visual tasks related to human poses and actions due to the lack of specialized…

Computer Vision and Pattern Recognition · Computer Science 2025-06-03 Dewen Zhang , Wangpeng An , Hayaru Shouno

Multi-modal intent detection aims to utilize various modalities to understand the user's intentions, which is essential for the deployment of dialogue systems in real-world scenarios. The two core challenges for multi-modal intent detection…

Computation and Language · Computer Science 2024-01-02 Shijue Huang , Libo Qin , Bingbing Wang , Geng Tu , Ruifeng Xu

This paper introduces a vision-based framework for capturing and understanding human behavior in industrial assembly lines, focusing on car door manufacturing. The framework leverages advanced computer vision techniques to estimate workers'…

Computer Vision and Pattern Recognition · Computer Science 2024-09-27 Konstantinos Papoutsakis , Nikolaos Bakalos , Konstantinos Fragkoulis , Athena Zacharia , Georgia Kapetadimitri , Maria Pateraki

We present IMPACT-HOI, a mixed-initiative framework for annotating egocentric procedural video by constructing structured event graphs for Human-Object Interactions (HOI), motivated by the need for high-quality structured supervision for…

‹ Prev 1 8 9 10 Next ›