中文
相关论文

相关论文: Inter-X: Towards Versatile Human-Human Interaction…

200 篇论文

Understanding human behavior from complementary egocentric (ego) and exocentric (exo) points of view enables the development of systems that can support workers in industrial environments and enhance their safety. However, progress in this…

Humans constantly interact with their surrounding environments. Current human-centric generative models mainly focus on synthesizing humans plausibly interacting with static scenes and objects, while the dynamic human action-reaction…

计算机视觉与模式识别 · 计算机科学 2024-03-19 Liang Xu , Yizhou Zhou , Yichao Yan , Xin Jin , Wenhan Zhu , Fengyun Rao , Xiaokang Yang , Wenjun Zeng

The lack of fine-grained joints (facial joints, hand fingers) is a fundamental performance bottleneck for state of the art skeleton action recognition models. Despite this bottleneck, community's efforts seem to be invested only in coming…

计算机视觉与模式识别 · 计算机科学 2021-11-25 Neel Trivedi , Anirudh Thatipelli , Ravi Kiran Sarvadevabhatla

The studies of human clothing for digital avatars have predominantly relied on synthetic datasets. While easy to collect, synthetic data often fall short in realism and fail to capture authentic clothing dynamics. Addressing this gap, we…

计算机视觉与模式识别 · 计算机科学 2024-04-30 Wenbo Wang , Hsuan-I Ho , Chen Guo , Boxiang Rong , Artur Grigorev , Jie Song , Juan Jose Zarate , Otmar Hilliges

To better interact with users, a social robot should understand the users' behavior, infer the intention, and respond appropriately. Machine learning is one way of implementing robot intelligence. It provides the ability to automatically…

机器人学 · 计算机科学 2022-11-01 Woo-Ri Ko , Minsu Jang , Jaeyeon Lee , Jaehong Kim

Humans intuitively understand that inanimate objects do not move by themselves, but that state changes are typically caused by human manipulation (e.g., the opening of a book). This is not yet the case for machines. In part this is because…

计算机视觉与模式识别 · 计算机科学 2023-04-25 Zicong Fan , Omid Taheri , Dimitrios Tzionas , Muhammed Kocabas , Manuel Kaufmann , Michael J. Black , Otmar Hilliges

In this paper, we introduce RealDex, a pioneering dataset capturing authentic dexterous hand grasping motions infused with human behavioral patterns, enriched by multi-view and multimodal visual data. Utilizing a teleoperation system, we…

In this paper, we introduce a new benchmark dataset named IPN Hand with sufficient size, variety, and real-world elements able to train and evaluate deep neural networks. This dataset contains more than 4,000 gesture samples and 800,000 RGB…

计算机视觉与模式识别 · 计算机科学 2020-10-21 Gibran Benitez-Garcia , Jesus Olivares-Mercado , Gabriel Sanchez-Perez , Keiji Yanai

Social interactions dominate our perceptions of the world and shape our daily behavior by attaching social meaning to acts as simple and spontaneous as gestures, facial expressions, voice, and speech. People mimic and otherwise respond to…

计算机视觉与模式识别 · 计算机科学 2026-04-27 Xiang Zhang , Xiaotian Li , Taoyue Wang , Nan Bi , Xin Zhou , Cody Zhou , Zoie Wang , Andrew Yang , Yuming Su , Jeff Cohn , Qiang Ji , Lijun Yin

Estimating the 3D structure of the human body from natural scenes is a fundamental aspect of visual perception. 3D human pose estimation is a vital step in advancing fields like AIGC and human-robot interaction, serving as a crucial…

计算机视觉与模式识别 · 计算机科学 2024-04-04 Jiong Wang , Fengyu Yang , Wenbo Gou , Bingliang Li , Danqi Yan , Ailing Zeng , Yijun Gao , Junle Wang , Yanqing Jing , Ruimao Zhang

We present Uni-Inter, a unified framework for human motion generation that supports a wide range of interaction scenarios: including human-human, human-object, and human-scene-within a single, task-agnostic architecture. In contrast to…

计算机视觉与模式识别 · 计算机科学 2025-11-18 Sheng Liu , Yuanzhi Liang , Jiepeng Wang , Sidan Du , Chi Zhang , Xuelong Li

Ambiguity is ubiquitous in human communication. Previous approaches in Human-Robot Interaction (HRI) have often relied on predefined interaction templates, leading to reduced performance in realistic and open-ended scenarios. To address…

机器人学 · 计算机科学 2023-10-19 Hanbo Zhang , Jie Xu , Yuchen Mo , Tao Kong

Micro-action is an imperceptible non-verbal behaviour characterised by low-intensity movement. It offers insights into the feelings and intentions of individuals and is important for human-oriented applications such as emotion recognition…

计算机视觉与模式识别 · 计算机科学 2024-06-04 Dan Guo , Kun Li , Bin Hu , Yan Zhang , Meng Wang

This paper introduces a new video-and-language dataset with human actions for multimodal logical inference, which focuses on intentional and aspectual expressions that describe dynamic human actions. The dataset consists of 200 videos,…

计算机视觉与模式识别 · 计算机科学 2021-06-29 Riko Suzuki , Hitomi Yanaka , Koji Mineshima , Daisuke Bekki

This paper introduces OmniMotion-X, a versatile multimodal framework for whole-body human motion generation, leveraging an autoregressive diffusion transformer in a unified sequence-to-sequence manner. OmniMotion-X efficiently supports…

计算机视觉与模式识别 · 计算机科学 2025-10-23 Guowei Xu , Yuxuan Bian , Ailing Zeng , Mingyi Shi , Shaoli Huang , Wen Li , Lixin Duan , Qiang Xu

We present Human Motions with Objects (HUMOTO), a high-fidelity dataset of human-object interactions for motion generation, computer vision, and robotics applications. Featuring 735 sequences (7,875 seconds at 30 fps), HUMOTO captures…

计算机视觉与模式识别 · 计算机科学 2025-10-16 Jiaxin Lu , Chun-Hao Paul Huang , Uttaran Bhattacharya , Qixing Huang , Yi Zhou

We present X-Avatar, a novel avatar model that captures the full expressiveness of digital humans to bring about life-like experiences in telepresence, AR/VR and beyond. Our method models bodies, hands, facial expressions and appearance in…

计算机视觉与模式识别 · 计算机科学 2023-03-10 Kaiyue Shen , Chen Guo , Manuel Kaufmann , Juan Jose Zarate , Julien Valentin , Jie Song , Otmar Hilliges

Human motion prediction aims to forecast future poses given a sequence of past 3D skeletons. While this problem has recently received increasing attention, it has mostly been tackled for single humans in isolation. In this paper, we explore…

计算机视觉与模式识别 · 计算机科学 2022-06-22 Wen Guo , Xiaoyu Bie , Xavier Alameda-Pineda , Francesc Moreno-Noguer

Building a socially intelligent agent involves many challenges, one of which is to teach the agent to speak guided by its value like a human. However, value-driven chatbots are still understudied in the area of dialogue systems. Most…

计算与语言 · 计算机科学 2022-07-25 Liang Qiu , Yizhou Zhao , Jinchao Li , Pan Lu , Baolin Peng , Jianfeng Gao , Song-Chun Zhu

We present a dataset with models of 14 articulated objects commonly found in human environments and with RGB-D video sequences and wrenches recorded of human interactions with them. The 358 interaction sequences total 67 minutes of human…

机器人学 · 计算机科学 2018-06-19 Roberto Martín-Martín , Clemens Eppner , Oliver Brock