中文
相关论文

相关论文: PGformer: Proxy-Bridged Game Transformer for Multi…

200 篇论文

Pedestrian behavior prediction is one of the major challenges for intelligent driving systems in urban environments. Pedestrians often exhibit a wide range of behaviors and adequate interpretations of those depend on various sources of…

计算机视觉与模式识别 · 计算机科学 2020-12-02 Amir Rasouli , Tiffany Yau , Mohsen Rohani , Jun Luo

We address the task of identifying distracted driving by analyzing in-car videos using efficient transformers. Although transformer models have achieved outstanding performance in human action recognition tasks, their high computational…

计算机视觉与模式识别 · 计算机科学 2026-03-03 Ricardo Pizarro , Roberto Valle , Rafael Barea , Jose M. Buenaposada , Luis Baumela , Luis Miguel Bergasa

Interacting with real-world objects in Mixed Reality (MR) often proves difficult when they are crowded, distant, or partially occluded, hindering straightforward selection and manipulation. We observe that these difficulties stem from…

人机交互 · 计算机科学 2025-07-25 Xiaoan Liu , Difan Jia , Xianhao Carton Liu , Mar Gonzalez-Franco , Chen Zhu-Tian

The real-time crash likelihood prediction model is an essential component of the proactive traffic safety management system. Over the years, numerous studies have attempted to construct a crash likelihood prediction model in order to…

机器学习 · 计算机科学 2023-08-30 B M Tazbiul Hassan Anik , Zubayer Islam , Mohamed Abdel-Aty

Proactive human-robot interaction (HRI) allows the receptionist robots to actively greet people and offer services based on vision, which has been found to improve acceptability and customer satisfaction. Existing approaches are either…

机器人学 · 计算机科学 2021-03-25 Yang Xue , Fan Wang , Hao Tian , Min Zhao , Jiangyong Li , Haiqing Pan , Yueqiang Dong

We propose to forecast future hand-object interactions given an egocentric video. Instead of predicting action labels or pixels, we directly predict the hand motion trajectory and the future contact points on the next active object (i.e.,…

计算机视觉与模式识别 · 计算机科学 2022-04-05 Shaowei Liu , Subarna Tripathi , Somdeb Majumdar , Xiaolong Wang

3D human-object interaction (HOI) anticipation aims to predict the future motion of humans and their manipulated objects, conditioned on the historical context. Generally, the articulated humans and rigid objects exhibit different motion…

计算机视觉与模式识别 · 计算机科学 2025-08-12 Xiaotong Lin , Tianming Liang , Jian-Fang Hu , Kun-Yu Lin , Yulei Kang , Chunwei Tian , Jianhuang Lai , Wei-Shi Zheng

Deictic gestures, like pointing, are a fundamental form of non-verbal communication, enabling humans to direct attention to specific objects or locations. This capability is essential in Human-Robot Interaction (HRI), where robots should be…

机器人学 · 计算机科学 2025-09-08 Luca Müller , Hassan Ali , Philipp Allgeuer , Lukáš Gajdošech , Stefan Wermter

In multiagent environments, several decision-making individuals interact while adhering to the dynamics constraints imposed by the environment. These interactions, combined with the potential stochasticity of the agents' decision-making…

Motion prediction is crucial in enabling safe motion planning for autonomous vehicles in interactive scenarios. It allows the planner to identify potential conflicts with other traffic agents and generate safe plans. Existing motion…

机器人学 · 计算机科学 2022-11-04 Qiao Sun , Xin Huang , Brian C. Williams , Hang Zhao

This report describes the 2nd place solution to the ECCV 2022 Human Body, Hands, and Activities (HBHA) from Egocentric and Multi-view Cameras Challenge: Action Recognition. This challenge aims to recognize hand-object interaction in an…

计算机视觉与模式识别 · 计算机科学 2022-10-21 Hoseong Cho , Seungryul Baek

Accurate trajectory prediction of nearby vehicles is crucial for the safe motion planning of automated vehicles in dynamic driving scenarios such as highway merging. Existing methods cannot initiate prediction for a vehicle unless observed…

机器人学 · 计算机科学 2023-06-12 Sajjad Mozaffari , Mreza Alipour Sormoli , Konstantinos Koufos , Graham Lee , Mehrdad Dianati

Predicting pedestrian motion trajectories is critical for path planning and motion control of autonomous vehicles. However, accurately forecasting crowd trajectories remains a challenging task due to the inherently multimodal and uncertain…

计算机视觉与模式识别 · 计算机科学 2025-08-07 Yu Liu , Zhijie Liu , Xiao Ren , You-Fu Li , He Kong

Despite recent progress, text-to-image models still struggle to generate semantically diverse and compositionally accurate multi-person interaction scenes, often collapsing to repetitive layouts, stereotypical poses, and poorly grounded…

计算机视觉与模式识别 · 计算机科学 2026-05-25 Wenxuan Peng , Bharath Hariharan , Hadar Averbuch-Elor

We present a GAN-based Transformer for general action-conditioned 3D human motion generation, including not only single-person actions but also multi-person interactive actions. Our approach consists of a powerful Action-conditioned motion…

计算机视觉与模式识别 · 计算机科学 2022-11-24 Liang Xu , Ziyang Song , Dongliang Wang , Jing Su , Zhicheng Fang , Chenjing Ding , Weihao Gan , Yichao Yan , Xin Jin , Xiaokang Yang , Wenjun Zeng , Wei Wu

As a fundamental aspect of human life, two-person interactions contain meaningful information about people's activities, relationships, and social settings. Human action recognition serves as the foundation for many smart applications, with…

计算机视觉与模式识别 · 计算机科学 2024-05-15 Yao Liu , Gangfeng Cui , Jiahui Luo , Xiaojun Chang , Lina Yao

Mapping discrete and dimensional models of emotion remains a persistent challenge in affective science and computing. This incompatibility hinders the combination of valuable data sets, creating a significant bottleneck for training robust…

人机交互 · 计算机科学 2025-11-18 Michal R. Wrobel

This paper proposes an interaction reasoning network for modelling spatio-temporal relationships between hands and objects in video. The proposed interaction unit utilises a Transformer module to reason about each acting hand, and its…

计算机视觉与模式识别 · 计算机科学 2022-01-14 Jian Ma , Dima Damen

The Transformer is a sequence model that forgoes traditional recurrent architectures in favor of a fully attention-based approach. Besides improving performance, an advantage of using attention is that it can also help to interpret a model…

人机交互 · 计算机科学 2019-06-14 Jesse Vig

Accurately predicting pedestrian motion is crucial for safe and reliable autonomous driving in complex urban environments. In this work, we present a 3D vehicle-conditioned pedestrian pose forecasting framework that explicitly incorporates…

计算机视觉与模式识别 · 计算机科学 2026-02-10 Guangxun Zhu , Xuan Liu , Nicolas Pugeault , Chongfeng Wei , Edmond S. L. Ho