English
Related papers

Related papers: Mask and Compress: Efficient Skeleton-based Action…

200 papers

This paper presents the ARN-LSTM architecture, a novel multi-stream action recognition model designed to address the challenge of simultaneously capturing spatial motion and temporal dynamics in action sequences. Traditional methods often…

Computer Vision and Pattern Recognition · Computer Science 2024-12-02 Chuanchuan Wang , Ahmad Sufril Azlan Mohmamed , Mohd Halim Bin Mohd Noor , Xiao Yang , Feifan Yi , Xiang Li

Human activity recognition (HAR) by wearable sensor devices embedded in the Internet of things (IOT) can play a significant role in remote health monitoring and emergency notification, to provide healthcare of higher standards. The purpose…

Machine Learning · Computer Science 2022-01-24 M. Abid , A. Khabou , Y. Ouakrim , H. Watel , S. Chemkhi , A. Mitiche , A. Benazza-Benyahia , N. Mezghani

Self-supervised pre-training paradigms have been extensively explored in the field of skeleton-based action recognition. In particular, methods based on masked prediction have pushed the performance of pre-training to a new height. However,…

Computer Vision and Pattern Recognition · Computer Science 2024-01-03 Ruizhuo Xu , Linzhi Huang , Mei Wang , Jiani Hu , Weihong Deng

Automatic facial expression recognition (FER) has gained much attention due to its applications in human-computer interaction. Among the approaches to improve FER tasks, this paper focuses on deep architecture with the attention mechanism.…

Computer Vision and Pattern Recognition · Computer Science 2026-03-09 Luan Pham , The Huynh Vu , Tuan Anh Tran

In this paper, we propose a coupled spatial-temporal attention (CSTA) model for skeleton-based action recognition, which aims to figure out the most discriminative joints and frames in spatial and temporal domains simultaneously.…

Computer Vision and Pattern Recognition · Computer Science 2019-09-24 Jiayun Wang

Pooling is a crucial operation in computer vision, yet the unique structure of skeletons hinders the application of existing pooling strategies to skeleton graph modelling. In this paper, we propose an Improved Graph Pooling Network,…

Computer Vision and Pattern Recognition · Computer Science 2024-04-26 Cong Wu , Xiao-Jun Wu , Tianyang Xu , Josef Kittler

In this paper, we tackle the problem of action recognition using body skeletons extracted from video sequences. Our approach lies in the continuity of recent works representing video frames by Gramian matrices that describe a trajectory on…

Computer Vision and Pattern Recognition · Computer Science 2019-09-10 Benjamin Szczapa , Mohamed Daoudi , Stefano Berretti , Alberto Del Bimbo , Pietro Pala , Estelle Massart

Contrastive learning has achieved great success in skeleton-based action recognition. However, most existing approaches encode the skeleton sequences as entangled spatiotemporal representations and confine the contrasts to the same level of…

Computer Vision and Pattern Recognition · Computer Science 2023-09-13 Cong Wu , Xiao-Jun Wu , Josef Kittler , Tianyang Xu , Sara Atito , Muhammad Awais , Zhenhua Feng

Graph convolutional networks (GCNs), which can model the human body skeletons as spatial and temporal graphs, have shown remarkable potential in skeleton-based action recognition. However, in the existing GCN-based methods, graph-structured…

Computer Vision and Pattern Recognition · Computer Science 2022-10-13 Han Chen , Yifan Jiang , Hanseok Ko

As the use of collaborative robots (cobots) in industrial manufacturing continues to grow, human action recognition for effective human-robot collaboration becomes increasingly important. This ability is crucial for cobots to act…

Computer Vision and Pattern Recognition · Computer Science 2023-06-12 Dustin Aganian , Mona Köhler , Sebastian Baake , Markus Eisenbach , Horst-Michael Gross

Considering the instance-level discriminative ability, contrastive learning methods, including MoCo and SimCLR, have been adapted from the original image representation learning task to solve the self-supervised skeleton-based action…

Computer Vision and Pattern Recognition · Computer Science 2024-10-02 Mengyuan Liu , Hong Liu , Tianyu Guo

Skeleton-based temporal action segmentation is a fundamental yet challenging task, playing a crucial role in enabling intelligent systems to perceive and respond to human activities. While fully-supervised methods achieve satisfactory…

Computer Vision and Pattern Recognition · Computer Science 2026-03-09 Hongsong Wang , Yiqin Shen , Pengbo Yan , Jie Gui

This paper presents the first investigation into the use of fully automated deep learning framework for assessing neonatal postoperative pain. It specifically investigates the use of Bilinear Convolutional Neural Network (B-CNN) to extract…

Computer Vision and Pattern Recognition · Computer Science 2021-08-04 Md Sirajus Salekin , Ghada Zamzmi , Dmitry Goldgof , Rangachar Kasturi , Thao Ho , Yu Sun

This paper presents a 2D skeleton-based action segmentation method with applications in fine-grained human activity recognition. In contrast with state-of-the-art methods which directly take sequences of 3D skeleton coordinates as inputs…

Computer Vision and Pattern Recognition · Computer Science 2024-04-29 Syed Waleed Hyder , Muhammad Usama , Anas Zafar , Muhammad Naufil , Fawad Javed Fateh , Andrey Konin , M. Zeeshan Zia , Quoc-Huy Tran

Skeleton-based human action recognition (HAR) has achieved remarkable progress with graph-based architectures. However, most existing methods remain body-centric, focusing on large-scale motions while neglecting subtle hand articulations…

Computer Vision and Pattern Recognition · Computer Science 2026-01-05 Seungyeon Cho , Tae-kyun Kim

In this paper, a contrastive representation learning framework is proposed to enhance human action segmentation via pre-training using trimmed (single action) skeleton sequences. Unlike previous representation learning works that are…

Computer Vision and Pattern Recognition · Computer Science 2025-09-09 Haitao Tian , Pierre Payeur

Most action recognition models treat human activities as unitary events. However, human activities often follow a certain hierarchy. In fact, many human activities are compositional. Also, these actions are mostly human-object interactions.…

Computer Vision and Pattern Recognition · Computer Science 2022-04-21 Mohammed Guermal , Rui Dai , Francois Bremond

Graph Convolutional Networks (GCNs), which model skeleton data as graphs, have obtained remarkable performance for skeleton-based action recognition. Particularly, the temporal dynamic of skeleton sequence conveys significant information in…

Computer Vision and Pattern Recognition · Computer Science 2020-12-17 Jianan Li , Xuemei Xie , Zhifu Zhao , Yuhan Cao , Qingzhe Pan , Guangming Shi

We present a module that extends the temporal graph of a graph convolutional network (GCN) for action recognition with a sequence of skeletons. Existing methods attempt to represent a more appropriate spatial graph on an intra-frame, but…

Computer Vision and Pattern Recognition · Computer Science 2020-10-20 Yuya Obinata , Takuma Yamamoto

The self-supervised pretraining paradigm has achieved great success in learning 3D action representations for skeleton-based action recognition using contrastive learning. However, learning effective representations for skeleton-based…

Computer Vision and Pattern Recognition · Computer Science 2026-05-06 Qiushuo Cheng , Jingjing Liu , Catherine Morgan , Alan Whone , Majid Mirmehdi
‹ Prev 1 8 9 10 Next ›