中文
相关论文

相关论文: SignAvatar: Sign Language 3D Motion Reconstruction…

200 篇论文

Sign Language Recognition (SLR) has garnered significant attention from researchers in recent years, particularly the intricate domain of Continuous Sign Language Recognition (CSLR), which presents heightened complexity compared to Isolated…

计算机视觉与模式识别 · 计算机科学 2024-02-23 Razieh Rastgoo , Kourosh Kiani , Sergio Escalera

Text-to-Avatar generation has recently made significant strides due to advancements in diffusion models. However, most existing work remains constrained by limited diversity, producing avatars with subtle differences in appearance for a…

计算机视觉与模式识别 · 计算机科学 2024-02-28 Weijing Tao , Biwen Lei , Kunhao Liu , Shijian Lu , Miaomiao Cui , Xuansong Xie , Chunyan Miao

We present a novel approach for generating animatable 3D-aware art avatars from a single image, with controllable facial expressions, head poses, and shoulder movements. Unlike previous reenactment methods, our approach utilizes a…

计算机视觉与模式识别 · 计算机科学 2024-03-27 Shaoxu Li

One of the factors that have hindered progress in the areas of sign language recognition, translation, and production is the absence of large annotated datasets. Towards this end, we introduce How2Sign, a multimodal and multiview continuous…

计算机视觉与模式识别 · 计算机科学 2021-04-02 Amanda Duarte , Shruti Palaskar , Lucas Ventura , Deepti Ghadiyaram , Kenneth DeHaan , Florian Metze , Jordi Torres , Xavier Giro-i-Nieto

Skeleton-based isolated sign language recognition (ISLR) demands fine-grained understanding of articulated motion across multiple spatial scales, from subtle finger movements to global body dynamics. Existing approaches typically rely on…

计算机视觉与模式识别 · 计算机科学 2026-05-13 Muxin Pu , Mei Kuan Lim , Chun Yong Chong , Chen Change Loy

Sign language is the window for people differently-abled to express their feelings as well as emotions. However, it remains challenging for people to learn sign language in a short time. To address this real-world challenge, in this work,…

计算机视觉与模式识别 · 计算机科学 2022-07-11 Yucheng Suo , Zhedong Zheng , Xiaohan Wang , Bang Zhang , Yi Yang

Sign language visual recognition from continuous multi-modal streams is still one of the most challenging fields. Recent advances in human actions recognition are exploiting the ascension of GPU-based learning from massive data, and are…

计算机视觉与模式识别 · 计算机科学 2020-09-23 Bassem Seddik , Najoua Essoukri Ben Amara

Recent advances in auto-regressive transformers have achieved remarkable success in generative modeling. However, text-to-3D generation remains challenging, primarily due to bottlenecks in learning discrete 3D representations. Specifically,…

计算机视觉与模式识别 · 计算机科学 2026-02-17 Zongcheng Han , Dongyan Cao , Haoran Sun , Yu Hong

Sign language production requires more than hand motion generation. Non-manual features, including mouthings, eyebrow raises, gaze, and head movements, are grammatically obligatory and cannot be recovered from manual articulators alone.…

计算机视觉与模式识别 · 计算机科学 2026-03-26 Alexandre Symeonidis-Herzig , Jianhe Low , Ozge Mercanoglu Sincan , Richard Bowden

Sign languages are dynamic visual languages that involve hand gestures, in combination with non manual elements such as facial expressions. While video recordings of sign language are commonly used for education and documentation, the…

计算机视觉与模式识别 · 计算机科学 2025-04-16 Janna Bruner , Amit Moryossef , Lior Wolf

Existing single-image 3D human avatar methods primarily rely on rigid joint transformations, limiting their ability to model realistic cloth dynamics. We present DynaAvatar, a zero-shot framework that reconstructs animatable 3D human…

计算机视觉与模式识别 · 计算机科学 2026-03-17 Joohyun Kwon , Geonhee Sim , Gyeongsik Moon

We present FlexAvatar, a flexible large reconstruction model for high-fidelity 3D head avatars with detailed dynamic deformation from single or sparse images, without requiring camera poses or expression labels. It leverages a…

计算机视觉与模式识别 · 计算机科学 2025-12-22 Cheng Peng , Zhuo Su , Liao Wang , Chen Guo , Zhaohu Li , Chengjiang Long , Zheng Lv , Jingxiang Sun , Chenyangguang Zhang , Yebin Liu

Speech-driven 3D facial animation aims to generate realistic lip movements and facial expressions for 3D head models from arbitrary audio clips. Although existing diffusion-based methods are capable of producing natural motions, their slow…

计算机视觉与模式识别 · 计算机科学 2025-09-05 Xuangeng Chu , Nabarun Goswami , Ziteng Cui , Hanqin Wang , Tatsuya Harada

In social robotics, endowing humanoid robots with the ability to generate bodily expressions of affect can improve human-robot interaction and collaboration, since humans attribute, and perhaps subconsciously anticipate, such traces to…

机器人学 · 计算机科学 2022-05-03 Mina Marmpena , Fernando Garcia , Angelica Lim , Nikolas Hemion , Thomas Wennekers

Generating talking head videos through a face image and a piece of speech audio still contains many challenges. ie, unnatural head movement, distorted expression, and identity modification. We argue that these issues are mainly because of…

计算机视觉与模式识别 · 计算机科学 2023-03-14 Wenxuan Zhang , Xiaodong Cun , Xuan Wang , Yong Zhang , Xi Shen , Yu Guo , Ying Shan , Fei Wang

Voice-controlled personal and home assistants (such as the Amazon Echo and Apple Siri) are becoming increasingly popular for a variety of applications. However, the benefits of these technologies are not readily accessible to Deaf or…

机器学习 · 计算机科学 2019-09-26 Al Amin Hosain , Panneer Selvam Santhalingam , Parth Pathak , Jana Kosecka , Huzefa Rangwala

Sign language, which conveys meaning through gestures, is the chief means of communication among deaf people. Recognizing sign language in natural settings presents significant challenges due to factors such as lighting, background clutter,…

计算机视觉与模式识别 · 计算机科学 2023-08-25 Bowen Shi

Try to generate new bridge types using generative artificial intelligence technology. The grayscale images of the bridge facade with the change of component width was rendered by 3dsMax animation software, and then the OpenCV module…

机器学习 · 计算机科学 2024-01-02 Hongjun Zhang

Subtle hand differences make sign language recognition challenging, yet many existing methods rely on encoders pretrained on generic action datasets that poorly capture such fine-grained cues. We propose a self-supervised pretraining method…

计算机视觉与模式识别 · 计算机科学 2026-05-05 Kunyuan Xie , Zhixi Cai , Kalin Stefanov

Existing approaches to animatable NeRF-based head avatars are either built upon face templates or use the expression coefficients of templates as the driving signal. Despite the promising progress, their performances are heavily bound by…

计算机视觉与模式识别 · 计算机科学 2023-05-04 Yuelang Xu , Hongwen Zhang , Lizhen Wang , Xiaochen Zhao , Han Huang , Guojun Qi , Yebin Liu