中文
相关论文

相关论文: Hierarchical I3D for Sign Spotting

200 篇论文

Sign language recognition is a challenging and often underestimated problem comprising multi-modal articulators (handshape, orientation, movement, upper body and face) that integrate asynchronously on multiple streams. Learning powerful…

计算机视觉与模式识别 · 计算机科学 2019-11-22 Hamid Reza Vaezi Joze , Oscar Koller

Sign language recognition (SLR) is a machine learning task aiming to identify signs in videos. Due to the scarcity of annotated data, unsupervised methods like contrastive learning have become promising in this field. They learn meaningful…

计算机视觉与模式识别 · 计算机科学 2026-03-09 Ariel Basso Madjoukeng , Jérôme Fink , Pierre Poitier , Edith Belise Kenmogne , Benoit Frenay

We introduce the problem of zero-shot sign language recognition (ZSSLR), where the goal is to leverage models learned over the seen sign class examples to recognize the instances of unseen signs. To this end, we propose to utilize the…

计算机视觉与模式识别 · 计算机科学 2019-07-25 Yunus Can Bilge , Nazli Ikizler-Cinbis , Ramazan Gokberk Cinbis

This paper examines two aspects of the isolated sign language recognition (ISLR) task. First, although a certain number of datasets is available, the data for individual sign languages is limited. It poses the challenge of cross-language…

计算机视觉与模式识别 · 计算机科学 2025-11-19 Ilya Ovodov , Petr Surovtsev , Karina Kvanchiani , Alexander Kapitanov , Alexander Nagaev

Sign language to spoken language audio translation is important to connect the hearing- and speech-challenged humans with others. We consider sign language videos with isolated sign sequences rather than continuous grammatical signing. Such…

计算机视觉与模式识别 · 计算机科学 2025-10-10 Harsh Kavediya , Vighnesh Nayak , Bheeshm Sharma , Balamurugan Palaniappan

Independent Sign Language Recognition is a complex visual recognition problem that combines several challenging tasks of Computer Vision due to the necessity to exploit and fuse information from hand gestures, body features and facial…

计算机视觉与模式识别 · 计算机科学 2020-12-11 Agelos Kratimenos , Georgios Pavlakos , Petros Maragos

Research on continuous sign language recognition (CSLR) is essential to bridge the communication gap between deaf and hearing individuals. Numerous previous studies have trained their models using the connectionist temporal classification…

计算机视觉与模式识别 · 计算机科学 2024-09-24 Ronglai Zuo , Fangyun Wei , Brian Mak

We present the first active learning tool for fine-grained 3D part labeling, a problem which challenges even the most advanced deep learning (DL) methods due to the significant structural variations among the small and intricate parts. For…

计算机视觉与模式识别 · 计算机科学 2024-04-02 Fenggen Yu , Yiming Qian , Francisca Gil-Ureta , Brian Jackson , Eric Bennett , Hao Zhang

This paper aims for the language-based product image retrieval task. The majority of previous works have made significant progress by designing network structure, similarity measurement, and loss function. However, they typically perform…

计算机视觉与模式识别 · 计算机科学 2021-02-19 Zhe Ma , Fenghao Liu , Jianfeng Dong , Xiaoye Qu , Yuan He , Shouling Ji

Recent progress in fine-grained gesture and action classification, and machine translation, point to the possibility of automated sign language recognition becoming a reality. A key stumbling block in making progress towards this goal is a…

计算机视觉与模式识别 · 计算机科学 2021-10-14 Samuel Albanie , Gül Varol , Liliane Momeni , Triantafyllos Afouras , Joon Son Chung , Neil Fox , Andrew Zisserman

Instance Image-Goal Navigation (IIN) requires autonomous agents to identify and navigate to a target object or location depicted in a reference image captured from any viewpoint. While recent methods leverage powerful novel view synthesis…

计算机视觉与模式识别 · 计算机科学 2025-06-10 Yijie Deng , Shuaihang Yuan , Geeta Chandra Raju Bethala , Anthony Tzes , Yu-Shen Liu , Yi Fang

Sign language is the primary language for people with a hearing loss. Sign language recognition (SLR) is the automatic recognition of sign language, which represents a challenging problem for computers, though some progress has been made…

计算机视觉与模式识别 · 计算机科学 2021-03-10 Roman Töngi

Recent Vision Language Models (VLMs) have demonstrated strong performance across a wide range of multimodal reasoning tasks. This raises the question of whether such general-purpose models can also address specialized visual recognition…

计算机视觉与模式识别 · 计算机科学 2026-04-14 Vaclav Javorek , Jakub Honzik , Ivan Gruber , Tomas Zelezny , Marek Hruz

Sign Language Translation (SLT) attempts to convert sign language videos into spoken sentences. However, many existing methods struggle with the disparity between visual and textual representations during end-to-end learning. Gloss-based…

计算机视觉与模式识别 · 计算机科学 2025-07-10 Sobhan Asasi , Mohamed Ilyes Lakhal , Richard Bowden

Word-level sign language recognition (WSLR) is a fundamental task in sign language interpretation. It requires models to recognize isolated sign words from videos. However, annotating WSLR data needs expert knowledge, thus limiting WSLR…

计算机视觉与模式识别 · 计算机科学 2020-03-18 Dongxu Li , Xin Yu , Chenchen Xu , Lars Petersson , Hongdong Li

Isolated Sign Language Recognition (SLR) has mostly been applied on datasets containing signs executed slowly and clearly by a limited group of signers. In real-world scenarios, however, we are met with challenging visual conditions,…

计算机视觉与模式识别 · 计算机科学 2023-08-17 Mathieu De Coster , Ellen Rushe , Ruth Holmes , Anthony Ventresque , Joni Dambre

This paper tackles the problem of zero-shot sign language recognition (ZSSLR), where the goal is to leverage models learned over the seen sign classes to recognize the instances of unseen sign classes. In this context, readily available…

计算机视觉与模式识别 · 计算机科学 2022-01-19 Yunus Can Bilge , Ramazan Gokberk Cinbis , Nazli Ikizler-Cinbis

This paper employs a multimodal approach for continuous sign recognition by first using ML for detecting the start and end frames of signs in videos of American Sign Language (ASL) sentences, and then by recognizing the segmented signs. For…

计算机视觉与模式识别 · 计算机科学 2025-12-08 Mingyu Zhao , Zhanfu Yang , Yang Zhou , Zhaoyang Xia , Can Jin , Xiaoxiao He , Dimitris N. Metaxas

Sign Language Processing (SLP) provides a foundation for a more inclusive future in language technology; however, the field faces several significant challenges that must be addressed to achieve practical, real-world applications. This work…

计算机视觉与模式识别 · 计算机科学 2024-09-25 Oline Ranum , David R. Wessels , Gomer Otterspeer , Erik J. Bekkers , Floris Roelofsen , Jari I. Andersen

Automatic Sign Language (SL) recognition is an important task in the computer vision community. To build a robust SL recognition system, we need a considerable amount of data which is lacking particularly in Indian sign language (ISL). In…

计算机视觉与模式识别 · 计算机科学 2024-09-30 Suvajit Patra , Arkadip Maitra , Megha Tiwari , K. Kumaran , Swathy Prabhu , Swami Punyeshwarananda , Soumitra Samanta