中文
相关论文

相关论文: PHASE: PHysically-grounded Abstract Social Events …

200 篇论文

Humans perceive their visual environment by directing their eyes towards relevant objects. The deployment of visual attention depends substantially on the stimulus's properties, higher cognitive processes, and biases and constraints of the…

神经元与认知 · 定量生物学 2025-07-22 Thomas Fabian

The ability to adapt to physical actions and constraints in an environment is crucial for embodied agents (e.g., robots) to effectively collaborate with humans. Such physically grounded human-AI collaboration must account for the increased…

机器学习 · 计算机科学 2025-09-30 Xuhui Kang , Sung-Wook Lee , Haolin Liu , Yuyan Wang , Yen-Ling Kuo

Robot-to-human object handover is an essential skill for robot assistants, from serving drinks at home to passing surgical tools in the operating room. We expect robots to perform handover robustly -- to release the object only after a firm…

机器人学 · 计算机科学 2026-05-07 Linfeng Li , Lin Shao , David Hsu

Current evaluation protocols predominantly assess physical reasoning in stationary scenes, creating a gap in evaluating agents' abilities to interact with dynamic events. While contemporary methods allow agents to modify initial scene…

人工智能 · 计算机科学 2024-03-26 Shiqian Li , Kewen Wu , Chi Zhang , Yixin Zhu

This paper describes our research on AI agents embodied in visual, virtual or physical forms, enabling them to interact with both users and their environments. These agents, which include virtual avatars, wearable devices, and robots, are…

By dynamic planning, we refer to the ability of the human brain to infer and impose motor trajectories related to cognitive decisions. A recent paradigm, active inference, brings fundamental insights into the adaptation of biological…

人工智能 · 计算机科学 2024-11-13 Matteo Priorelli , Ivilin Peev Stoianov

This work presents and experimentally test the framework used by our context-aware, distributed team of small Unmanned Aerial Systems (SUAS) capable of operating in real-time, in an autonomous fashion, and under constrained communications.…

Part-level Action Parsing aims at part state parsing for boosting action recognition in videos. Despite of dramatic progresses in the area of video classification research, a severe problem faced by the community is that the detailed…

计算机视觉与模式识别 · 计算机科学 2021-11-08 Xuanhan Wang , Xiaojia Chen , Lianli Gao , Lechao Chen , Jingkuan Song

When working around other agents such as humans, it is important to model their perception capabilities to predict and make sense of their behavior. In this work, we consider agents whose perception capabilities are determined by their…

机器人学 · 计算机科学 2025-08-12 Maulik Bhatt , HongHao Zhen , Monroe Kennedy , Negar Mehr

Understanding the causal influence of one agent on another agent is crucial for safely deploying artificially intelligent systems such as automated vehicles and mobile robots into human-inhabited environments. Existing models of causal…

多智能体系统 · 计算机科学 2025-05-26 Ashwin George , Luciano Cavalcante Siebert , David A. Abbink , Arkady Zgonnikov

Human social behavior is structured by relationships. We form teams, groups, tribes, and alliances at all scales of human life. These structures guide multi-agent cooperation and competition, but when we observe others these underlying…

人工智能 · 计算机科学 2019-01-21 Michael Shum , Max Kleiman-Weiner , Michael L. Littman , Joshua B. Tenenbaum

Human perception is based on unconscious inference, where sensory input integrates with prior information. This phenomenon, known as context dependency, helps in facing the uncertainty of the external world with predictions built upon…

机器人学 · 计算机科学 2022-07-19 Carlo Mazzola , Francesco Rea , Alessandra Sciutti

When humans navigate a crowed space such as a university campus or the sidewalks of a busy street, they follow common sense rules based on social etiquette. In this paper, we argue that in order to enable the design of new algorithms that…

计算机视觉与模式识别 · 计算机科学 2016-01-07 Alexandre Robicquet , Alexandre Alahi , Amir Sadeghian , Bryan Anenberg , John Doherty , Eli Wu , Silvio Savarese

Integrating robots into populated environments is a complex challenge that requires an understanding of human social dynamics. In this work, we propose to model social motion forecasting in a shared human-robot representation space, which…

机器人学 · 计算机科学 2024-04-09 Esteve Valls Mascaro , Yashuai Yan , Dongheui Lee

Perceiving the surrounding environment in terms of objects is useful for any general purpose intelligent agent. In this paper, we investigate a fundamental mechanism making object perception possible, namely the identification of…

人工智能 · 计算机科学 2018-10-12 Nicolas Le Hir , Olivier Sigaud , Alban Laflaquière

We develop an approach for active semantic perception which refers to using the semantics of the scene for tasks such as exploration. We build a compact, hierarchical multi-layer scene graph that can represent large, complex indoor…

机器人学 · 计算机科学 2025-10-08 Huayi Tang , Pratik Chaudhari

Most human behaviors consist of multiple parts, steps, or subtasks. These structures guide our action planning and execution, but when we observe others, the latent structure of their actions is typically unobservable, and must be inferred…

人工智能 · 计算机科学 2018-09-28 Ryo Nakahashi , Chris L. Baker , Joshua B. Tenenbaum

Visual-based human action recognition can be found in various application fields, e.g., surveillance systems, sports analytics, medical assistive technologies, or human-robot interaction frameworks, and it concerns the identification and…

计算机视觉与模式识别 · 计算机科学 2024-12-23 Antonios Gasteratos , Stavros N. Moutsis , Konstantinos A. Tsintotas , Yiannis Aloimonos

Arguably, the visual perception of conversational agents to the physical world is a key way for them to exhibit the human-like intelligence. Image-grounded conversation is thus proposed to address this challenge. Existing works focus on…

计算与语言 · 计算机科学 2021-06-24 Zujie Liang , Huang Hu , Can Xu , Chongyang Tao , Xiubo Geng , Yining Chen , Fan Liang , Daxin Jiang

Believable proxies of human behavior can empower interactive applications ranging from immersive environments to rehearsal spaces for interpersonal communication to prototyping tools. In this paper, we introduce generative…