中文
相关论文

相关论文: Re-mine, Learn and Reason: Exploring the Cross-mod…

200 篇论文

Zero-shot Human-Object Interaction detection aims to localize humans and objects in an image and recognize their interaction, even when specific verb-object pairs are unseen during training. Recent works have shown promising results using…

计算机视觉与模式识别 · 计算机科学 2025-10-30 Chanhyeong Yang , Taehoon Song , Jihwan Park , Hyunwoo J. Kim

Existing hand-object interactions (HOI) methods are largely limited to rigid objects, while 4D reconstruction methods of articulated objects generally require pre-scanning the object or even multi-view videos. It remains an unexplored but…

计算机视觉与模式识别 · 计算机科学 2026-03-30 Zikai Wang , Zhilu Zhang , Yiqing Wang , Hui Li , Wangmeng Zuo

Generating realistic 3D human-object interactions (HOIs) remains a challenging task due to the difficulty of modeling detailed interaction dynamics. Existing methods treat human and object motions independently, resulting in physically…

计算机视觉与模式识别 · 计算机科学 2025-10-21 Lin Wu , Zhixiang Chen , Jianglin Lan

We present HOIGaze - a novel learning-based approach for gaze estimation during hand-object interactions (HOI) in extended reality (XR). HOIGaze addresses the challenging HOI setting by building on one key insight: The eye, hand, and head…

计算机视觉与模式识别 · 计算机科学 2025-04-29 Zhiming Hu , Daniel Haeufle , Syn Schmitt , Andreas Bulling

Human-Object Interaction (HOI) detection is a task to localize humans and objects in an image and predict the interactions in human-object pairs. In real-world scenarios, HOI detection models need systematic generalization, i.e.,…

计算机视觉与模式识别 · 计算机科学 2024-04-15 Kentaro Takemoto , Moyuru Yamada , Tomotake Sasaki , Hisanao Akima

Recent advances in deep neural networks have achieved significant progress in detecting individual objects from an image. However, object detection is not sufficient to fully understand a visual scene. Towards a deeper visual understanding,…

计算机视觉与模式识别 · 计算机科学 2023-12-21 Bumsoo Kim , Taeho Choi , Jaewoo Kang , Hyunwoo J. Kim

We propose CG-HOI, the first method to address the task of generating dynamic 3D human-object interactions (HOIs) from text. We model the motion of both human and object in an interdependent fashion, as semantically rich human motion rarely…

计算机视觉与模式识别 · 计算机科学 2024-05-20 Christian Diller , Angela Dai

Interaction is one of the core abilities of humanoid robots. However, most existing frameworks focus on non-interactive whole-body control, which limits their practical applicability. In this work, we develop InterReal, a unified…

机器人学 · 计算机科学 2026-03-10 Dayang Liang , Yuhang Lin , Xinzhe Liu , Jiyuan Shi , Yunlong Liu , Chenjia Bai

This paper revisits human-object interaction (HOI) recognition at image level without using supervisions of object location and human pose. We name it detection-free HOI recognition, in contrast to the existing detection-supervised…

计算机视觉与模式识别 · 计算机科学 2021-07-29 Ying Jin , Yinpeng Chen , Lijuan Wang , Jianfeng Wang , Pei Yu , Zicheng Liu , Jenq-Neng Hwang

Robots are becoming increasingly popular in a wide range of environments due to their exceptional work capacity, precision, efficiency, and scalability. This development has been further encouraged by advances in Artificial Intelligence,…

人机交互 · 计算机科学 2023-12-14 Daniel Weber

Generating realistic and physically plausible 3D Human-Object Interactions (HOI) remains a key challenge in motion generation. One primary reason is that describing these physical constraints with words alone is difficult. To address this…

计算机视觉与模式识别 · 计算机科学 2026-03-26 Songjin Cai , Linjie Zhong , Ling Guo , Changxing Ding

Human-Object Interaction (HOI) detection, which localizes and infers relationships between human and objects, plays an important role in scene understanding. Although two-stage HOI detectors have advantages of high efficiency in training…

计算机视觉与模式识别 · 计算机科学 2023-04-18 Jeeseung Park , Jin-Woo Park , Jong-Seok Lee

This paper explores a cross-modality synthesis task that infers 3D human-object interactions (HOIs) from a given text-based instruction. Existing text-to-HOI synthesis methods mainly deploy a direct mapping from texts to object-specific 3D…

计算机视觉与模式识别 · 计算机科学 2025-03-05 Xuehao Gao , Yang Yang , Shaoyi Du , Yang Wu , Yebin Liu , Guo-Jun Qi

Human-Machine Interaction (HMI) systems have gained huge interest in recent years, with reference expression comprehension being one of the main challenges. Traditionally human-machine interaction has been mostly limited to speech and…

人机交互 · 计算机科学 2023-06-21 Aman Jain , Anirudh Reddy Kondapally , Kentaro Yamada , Hitomi Yanaka

Resting-state functional magnetic resonance imaging (fMRI) has emerged as a cornerstone for psychiatric diagnosis, yet most approaches rely on pairwise brain cortical or sub-cortical connectivities that overlooks higher-order interactions…

机器学习 · 计算机科学 2026-04-21 Kunyu Zhang , Qiang Li , Vince D. Calhoun , Shujian Yu

Accurately modeling detailed interactions between human/hand and object is an appealing yet challenging task. Current multi-view capture systems are only capable of reconstructing multiple subjects into a single, unified mesh, which fails…

计算机视觉与模式识别 · 计算机科学 2024-03-22 Jiajun Zhang , Yuxiang Zhang , Hongwen Zhang , Xiao Zhou , Boyao Zhou , Ruizhi Shao , Zonghai Hu , Yebin Liu

In many real-world imitation learning tasks, the demonstrator and the learner have to act under different observation spaces. This situation brings significant obstacles to existing imitation learning approaches, since most of them learn…

机器学习 · 计算机科学 2022-10-10 Xin-Qiang Cai , Yao-Xiang Ding , Zi-Xuan Chen , Yuan Jiang , Masashi Sugiyama , Zhi-Hua Zhou

Understanding humans from LiDAR point clouds is one of the most critical tasks in autonomous driving due to its close relationships with pedestrian safety, yet it remains challenging in the presence of diverse human-object interactions and…

计算机视觉与模式识别 · 计算机科学 2026-03-18 Daniel Sungho Jung , Dohee Cho , Kyoung Mu Lee

Multimodal intent recognition aims to infer human intents by jointly modeling various modalities, playing a pivotal role in real-world dialogue systems. However, current methods struggle to model hierarchical semantics underlying complex…

多媒体 · 计算机科学 2026-03-05 Qianrui Zhou , Hua Xu , Yunjin Gu , Yifan Wang , Songze Li , Hanlei Zhang

Resolving real-world human-object interactions in images is a many-to-many challenge, in which disentangling fine-grained concurrent physical contact is particularly difficult. Existing semantic contact estimation methods are either limited…

计算机视觉与模式识别 · 计算机科学 2026-05-13 Sravan Chittupalli , Ayush Jain , Dong Huang
‹ 上一页 1 8 9 10 下一页 ›