中文
相关论文

相关论文: FT-HID: A Large Scale RGB-D Dataset for First and …

200 篇论文

Recognizing pain in video is crucial for improving patient-computer interaction systems, yet traditional data collection in this domain raises significant ethical and logistical challenges. This study introduces a novel approach that…

计算机视觉与模式识别 · 计算机科学 2024-09-26 Jonas Nasimzada , Jens Kleesiek , Ken Herrmann , Alina Roitberg , Constantin Seibold

We present HOI4D, a large-scale 4D egocentric dataset with rich annotations, to catalyze the research of category-level human-object interaction. HOI4D consists of 2.4M RGB-D egocentric video frames over 4000 sequences collected by 4…

计算机视觉与模式识别 · 计算机科学 2024-01-04 Yunze Liu , Yun Liu , Che Jiang , Kangbo Lyu , Weikang Wan , Hao Shen , Boqiang Liang , Zhoujie Fu , He Wang , Li Yi

Humans constantly interact with daily objects to accomplish tasks. To understand such interactions, computers need to reconstruct these from cameras observing whole-body interaction with scenes. This is challenging due to occlusion between…

计算机视觉与模式识别 · 计算机科学 2022-10-04 Yinghao Huang , Omid Tehari , Michael J. Black , Dimitrios Tzionas

To fluently collaborate with people, robots need the ability to recognize human activities accurately. Although modern robots are equipped with various sensors, robust human activity recognition (HAR) still remains a challenging task for…

机器人学 · 计算机科学 2020-08-17 Md Mofijul Islam , Tariq Iqbal

Learning the prior knowledge of the 3D human-object spatial relation is crucial for reconstructing human-object interaction from images and understanding how humans interact with objects in 3D space. Previous works learn this prior from…

计算机视觉与模式识别 · 计算机科学 2024-08-01 Chaofan Huo , Ye Shi , Jingya Wang

Data visualizations are powerful tools for communicating patterns in quantitative data. Yet understanding any data visualization is no small feat -- succeeding requires jointly making sense of visual, numerical, and linguistic inputs…

人机交互 · 计算机科学 2025-05-26 Arnav Verma , Kushin Mukherjee , Christopher Potts , Elisa Kreiss , Judith E. Fan

Large-scale is a trend in person re-identification (re-id). It is important that real-time search be performed in a large gallery. While previous methods mostly focus on discriminative learning, this paper makes the attempt in integrating…

计算机视觉与模式识别 · 计算机科学 2017-05-08 Fuqing Zhu , Xiangwei Kong , Liang Zheng , Haiyan Fu , Qi Tian

Thanks for the cross-modal retrieval techniques, visible-infrared (RGB-IR) person re-identification (Re-ID) is achieved by projecting them into a common space, allowing person Re-ID in 24-hour surveillance systems. However, with respect to…

计算机视觉与模式识别 · 计算机科学 2022-08-05 Xinyu Lin , Jinxing Li , Zeyu Ma , Huafeng Li , Shuang Li , Kaixiong Xu , Guangming Lu , David Zhang

Human Sensing, a field that leverages technology to monitor human activities, psycho-physiological states, and interactions with the environment, enhances our understanding of human behavior and drives the development of advanced services…

Human affect recognition has been a significant topic in psychophysics and computer vision. However, the currently published datasets have many limitations. For example, most datasets contain frames that contain only information about…

计算机视觉与模式识别 · 计算机科学 2023-09-18 Zhihang Ren , Jefferson Ortega , Yifan Wang , Zhimin Chen , Yunhui Guo , Stella X. Yu , David Whitney

Skeleton-based human action recognition has been drawing more interest recently due to its low sensitivity to appearance changes and the accessibility of more skeleton data. However, even the 3D skeletons captured in practice are still…

计算机视觉与模式识别 · 计算机科学 2022-09-26 Cunling Bian , Wei Feng , Fanbo Meng , Song Wang

Person re-identification (Re-ID) is one of the primary components of an automated visual surveillance system. It aims to automatically identify/search persons in a multi-camera network having non-overlapping field-of-views. Owing to its…

计算机视觉与模式识别 · 计算机科学 2022-03-01 Asmat Zahra , Nazia Perwaiz , Muhammad Shahzad , Muhammad Moazam Fraz

Text-Based Person Search (TBPS) has seen significant progress with vision-language models (VLMs), yet it remains constrained by limited training data and the fact that VLMs are not inherently pre-trained for pedestrian-centric recognition.…

计算机视觉与模式识别 · 计算机科学 2026-01-22 Nilanjana Chatterjee , Sidharatha Garg , A V Subramanyam , Brejesh Lall

We present a comprehensive framework for egocentric interaction recognition using markerless 3D annotations of two hands manipulating objects. To this end, we propose a method to create a unified dataset for egocentric 3D interaction…

计算机视觉与模式识别 · 计算机科学 2021-08-25 Taein Kwon , Bugra Tekin , Jan Stuhmer , Federica Bogo , Marc Pollefeys

Multimodal fusion frameworks for Human Action Recognition (HAR) using depth and inertial sensor data have been proposed over the years. In most of the existing works, fusion is performed at a single level (feature level or decision level),…

机器学习 · 计算机科学 2019-10-28 Zeeshan Ahmad , Naimul Khan

Training embodied agents to understand 3D scenes as humans do requires large-scale data of people meaningfully interacting with diverse environments, yet such data is scarce. Real-world capture is costly and limited to controlled settings,…

计算机视觉与模式识别 · 计算机科学 2026-05-27 Nikita Kister , Pradyumna YM , István Sárándi , Jiayi Wang , Anna Khoreva , Gerard Pons-Moll

Humans learn to imitate by observing others. However, robot imitation learning generally requires expert demonstrations in the first-person view (FPV). Collecting such FPV videos for every robot could be very expensive. Third-person…

机器人学 · 计算机科学 2021-08-03 Jinghuan Shang , Michael S. Ryoo

We present PANDA, the first gigaPixel-level humAN-centric viDeo dAtaset, for large-scale, long-term, and multi-object visual analysis. The videos in PANDA were captured by a gigapixel camera and cover real-world scenes with both wide…

计算机视觉与模式识别 · 计算机科学 2020-03-11 Xueyang Wang , Xiya Zhang , Yinheng Zhu , Yuchen Guo , Xiaoyun Yuan , Liuyu Xiang , Zerun Wang , Guiguang Ding , David J Brady , Qionghai Dai , Lu Fang

Due to the high complexity and occlusion, insufficient perception in the crowded urban intersection can be a serious safety risk for both human drivers and autonomous algorithms, whereas CVIS (Cooperative Vehicle Infrastructure System) is a…

计算机视觉与模式识别 · 计算机科学 2021-06-08 Huanan Wang , Xinyu Zhang , Jun Li , Zhiwei Li , Lei Yang , Shuyue Pan , Yongqiang Deng

The primary objective of the dataset is to provide a better understanding of the coupling between human actions and gaze in a shared working environment with a cobot, with the aim of signifcantly enhancing the effciency and safety of…

机器人学 · 计算机科学 2025-03-17 Maxence Grand , Damien Pellier , Francis Jambon
‹ 上一页 1 8 9 10 下一页 ›