English
Related papers

Related papers: PoseBridge: Bridging the Skeletonization Gap for Z…

200 papers

Skeletal Action recognition from an egocentric view is important for applications such as interfaces in AR/VR glasses and human-robot interaction, where the device has limited resources. Most of the existing skeletal action recognition…

Computer Vision and Pattern Recognition · Computer Science 2023-09-20 Junan Lin , Zhichao Sun , Enjie Cao , Taein Kwon , Mahdi Rad , Marc Pollefeys

Zero-shot skeleton-based action recognition aims to classify unseen skeleton-based human actions without prior exposure to such categories during training. This task is extremely challenging due to the difficulty in generalizing from known…

Computer Vision and Pattern Recognition · Computer Science 2025-07-25 Kai Zhou , Shuhai Zhang , Zeng You , Jinwu Hu , Mingkui Tan , Fei Liu

Visual neural decoding seeks to reconstruct or infer perceived visual stimuli from brain activity patterns, providing critical insights into human cognition and enabling transformative applications in brain-computer interfaces and…

Computer Vision and Pattern Recognition · Computer Science 2025-11-11 Wenjiang Zhang , Sifeng Wang , Yuwei Su , Xinyu Li , Chen Zhang , Suyu Zhong

Different from traditional video retrieval, sign language retrieval is more biased towards understanding the semantic information of human actions contained in video clips. Previous works typically only encode RGB videos to obtain…

Computer Vision and Pattern Recognition · Computer Science 2024-07-24 Longtao Jiang , Min Wang , Zecheng Li , Yao Fang , Wengang Zhou , Houqiang Li

Robustness to domain changes is a key capability for effective deployment of human action recognition systems in real-world scenarios, where action categories at inference can present important domain shifts or even unseen actions from…

Computer Vision and Pattern Recognition · Computer Science 2026-05-22 Yannick Porto , Renato Martins , Thomas Chalumeau , Cedric Demonceaux

Zero-shot action recognition is challenging due to the semantic gap between seen and unseen classes. We present a novel framework that enhances CLIP with disentangled embeddings and semantic-guided interaction. A Motion Separation Module…

Computer Vision and Pattern Recognition · Computer Science 2026-04-21 Yiming Wang , Frederick W. B. Li , Jingyun Wang

Current state-of-the-art methods for skeleton-based action recognition are supervised and rely on labels. The reliance is limiting the performance due to the challenges involved in annotation and mislabeled data. Unsupervised methods have…

Computer Vision and Pattern Recognition · Computer Science 2020-12-09 Jingyuan Li , Eli Shlizerman

In zero-shot skeleton-based action recognition (ZSAR), aligning skeleton features with the text features of action labels is essential for accurately predicting unseen actions. ZSAR faces a fundamental challenge in bridging the modality gap…

Computer Vision and Pattern Recognition · Computer Science 2025-07-17 Jeonghyeok Do , Munchurl Kim

We introduce SynSE, a novel syntactically guided generative approach for Zero-Shot Learning (ZSL). Our end-to-end approach learns progressively refined generative embedding spaces constrained within and across the involved modalities…

Computer Vision and Pattern Recognition · Computer Science 2021-06-30 Pranay Gupta , Divyanshu Sharma , Ravi Kiran Sarvadevabhatla

Zero-shot skeleton action recognition is a non-trivial task that requires robust unseen generalization with prior knowledge from only seen classes and shared semantics. Existing methods typically build the skeleton-semantics interactions by…

Computer Vision and Pattern Recognition · Computer Science 2024-11-27 Yang Chen , Jingcai Guo , Song Guo , Dacheng Tao

Human action recognition (HAR) has achieved impressive results with deep learning models, but their decision-making process remains opaque due to their black-box nature. Ensuring interpretability is crucial, especially for real-world…

Computer Vision and Pattern Recognition · Computer Science 2025-04-18 Jongseo Lee , Wooil Lee , Gyeong-Moon Park , Seong Tae Kim , Jinwoo Choi

Word-level sign language recognition (WSLR) has attracted attention because it is expected to overcome the communication barrier between people with speech impairment and those who can hear. In the WSLR problem, a method designed for action…

Computer Vision and Pattern Recognition · Computer Science 2024-11-21 Mizuki Maruyama , Shrey Singh , Katsufumi Inoue , Partha Pratim Roy , Masakazu Iwamura , Michifumi Yoshioka

Zero-shot action recognition, which addresses the issue of scalability and generalization in action recognition and allows the models to adapt to new and unseen actions dynamically, is an important research topic in computer vision…

Computer Vision and Pattern Recognition · Computer Science 2025-08-25 Jidong Kuang , Hongsong Wang , Chaolei Han , Yang Zhang , Jie Gui

Sign Language Recognition (SLR) systems aim to be embedded in video stream platforms to recognize the sign performed in front of a camera. SLR research has taken advantage of recent advances in pose estimation models to use skeleton…

Computer Vision and Pattern Recognition · Computer Science 2023-04-13 David Laines , Gissella Bejarano , Miguel Gonzalez-Mendoza , Gilberto Ochoa-Ruiz

Human action recognition is pivotal in computer vision, with applications ranging from surveillance to human-robot interaction. Despite the effectiveness of supervised skeleton-based methods, their reliance on exhaustive annotation limits…

Computer Vision and Pattern Recognition · Computer Science 2026-05-08 Yuxi Zhou , Zhengbo Zhang , Jingyu Pan , Zhiyu Lin , Zhigang Tu

Sparse encoders offer high-precision retrieval by representing term importance within a vocabulary space, yet their English-centric structures pose a critical impediment to language transfer for non-English languages. To overcome this…

Information Retrieval · Computer Science 2026-05-26 Seongtae Hong , Youngjoon Jang , Jia-Heui Ju , Hyeonseok Moon , Heuiseok Lim

Zero-Shot Action Recognition (ZSAR) aims to recognize video actions that have never been seen during training. Most existing methods assume a shared semantic space between seen and unseen actions and intend to directly learn a mapping from…

Computer Vision and Pattern Recognition · Computer Science 2022-06-23 Zhiyi Gao , Yonghong Hou , Wanqing Li , Zihui Guo , Bin Yu

Zero-shot action recognition (ZSAR) requires collaborative multi-modal spatiotemporal understanding. However, finetuning CLIP directly for ZSAR yields suboptimal performance, given its inherent constraints in capturing essential temporal…

Computer Vision and Pattern Recognition · Computer Science 2025-02-11 Yating Yu , Congqi Cao , Yueran Zhang , Qinyi Lv , Lingtong Min , Yanning Zhang

Recently, pose-based action recognition has gained more and more attention due to the better performance compared with traditional appearance-based methods. However, there still exist two problems to be further solved. First, existing…

Computer Vision and Pattern Recognition · Computer Science 2018-05-23 Wei Wang , Jinjin Zhang , Chenyang Si , Liang Wang

Human skeleton, as a compact representation of human action, has received increasing attention in recent years. Many skeleton-based action recognition methods adopt graph convolutional networks (GCN) to extract features on top of human…

Computer Vision and Pattern Recognition · Computer Science 2022-04-05 Haodong Duan , Yue Zhao , Kai Chen , Dahua Lin , Bo Dai