中文
相关论文

相关论文: RingGesture: A Ring-Based Mid-Air Gesture Typing S…

200 篇论文

In this paper, we propose a novel touchless typing interface that makes use of an on-screen QWERTY keyboard and a smartphone camera. The keyboard was divided into nine color-coded clusters. The user moved their head toward clusters, which…

人机交互 · 计算机科学 2020-10-13 Shivam Rustagi , Aakash Garg , Pranay Raj Anand , Rajesh Kumar , Yaman Kumar , Rajiv Ratn Shah

Hand Gesture Recognition (HGR) enables intuitive human-computer interactions in various real-world contexts. However, existing frameworks often struggle to meet the real-time requirements essential for practical HGR applications. This study…

计算机视觉与模式识别 · 计算机科学 2026-04-17 Oluwaleke Yusuf , Maki Habib , Mohamed Moustafa

Millimeter wave (mmWave) radar sensors play a vital role in hand gesture recognition (HGR) by detecting subtle motions while preserving user privacy. However, the limited scale of radar datasets hinders the performance. Existing synthetic…

人机交互 · 计算机科学 2025-04-24 Jiaqi Tang , Xinbo Xu , Yinsong Xu , Qingchao Chen

This paper introduces Wisture, a new online machine learning solution for recognizing touch-less dynamic hand gestures on a smartphone. Wisture relies on the standard Wi-Fi Received Signal Strength (RSS) using a Long Short-Term Memory…

人机交互 · 计算机科学 2017-08-01 Mohamed Abudulaziz Ali Haseeb , Ramviyas Parasuraman

In the fast-evolving field of Generative AI, platforms like MidJourney, DALL-E, and Stable Diffusion have transformed Text-to-Image (T2I) Generation. However, despite their impressive ability to create high-quality images, they often…

计算机视觉与模式识别 · 计算机科学 2024-09-19 Abhinaw Jagtap , Nachiket Tapas , R. G. Brajesh

In this paper, we introduce a new benchmark dataset for the challenging writing in the air (WiTA) task -- an elaborate task bridging vision and NLP. WiTA implements an intuitive and natural writing method with finger movement for…

计算机视觉与模式识别 · 计算机科学 2021-04-20 Ue-Hwan Kim , Yewon Hwang , Sun-Kyung Lee , Jong-Hwan Kim

Text-guided multispectral object detection uses text semantics to guide semantic-aware cross-modal interaction between RGB and IR for more robust perception. However, notable limitations remain: (1) existing methods often use text only as…

计算机视觉与模式识别 · 计算机科学 2026-04-14 Jiaqi Wu , Zhen Wang , Enhao Huang , Kangqing Shen , Yulin Wang , Yang Yue , Yifan Pu , Gao Huang

Gesture recognition is a cornerstone of Human-Computer Interaction (HCI) for smart eyewear, enabling natural and device-free control in augmented reality environments. Traditional vision-based approaches face significant challenges…

机器学习 · 计算机科学 2026-05-14 Pietro Bartoli , Christian Veronesi , Tommaso Bondini , Andrea Giudici , Franco Zappa

While the efficacy of deep learning models heavily relies on data, gathering and annotating data for specific tasks, particularly when addressing novel or sensitive subjects lacking relevant datasets, poses significant time and resource…

计算机视觉与模式识别 · 计算机科学 2025-06-25 Quang-Binh Nguyen , Trong-Vu Hoang , Ngoc-Do Tran , Tam V. Nguyen , Minh-Triet Tran , Trung-Nghia Le

Remote sensing imagery has attracted significant attention in recent years due to its instrumental role in global environmental monitoring, land usage monitoring, and more. As image databases grow each year, performing automatic…

计算机视觉与模式识别 · 计算机科学 2026-01-21 Jielu Zhang , Zhongliang Zhou , Gengchen Mai , Mengxuan Hu , Zihan Guan , Sheng Li , Lan Mu

Based on the DeepSORT algorithm, this study explores the application of visual tracking technology in intelligent human-computer interaction, especially in the field of gesture recognition and tracking. With the rapid development of…

人机交互 · 计算机科学 2025-05-13 Tong Zhang , Fenghua Shao , Runsheng Zhang , Yifan Zhuang , Liuqingqing Yang

Nowadays, Hearth Rate (HR) monitoring is a key feature of almost all wrist-worn devices exploiting photoplethysmography (PPG) sensors. However, arm movements affect the performance of PPG-based HR tracking. This issue is usually addressed…

信号处理 · 电气工程与系统科学 2022-10-21 Panagiotis Kasnesis , Lazaros Toumanidis , Alessio Burrello , Christos Chatzigeorgiou , Charalampos Z. Patrikakis

Computers today aren't just confined to laptops and desktops. Mobile gadgets like mobile phones and laptops also make use of it. However, one input device that hasn't changed in the last 50 years is the QWERTY keyboard. Users of virtual…

人机交互 · 计算机科学 2022-08-02 Pabbathi Sri Charan , Saksham Gupta , Satvik Agrawal , Gadupudi Sahithi Sindhu

This paper introduces the concept of augmented conversation, which aims to support co-located in-person conversations via embedded speech-driven on-the-fly referencing in augmented reality (AR). Today computing technologies like smartphones…

人机交互 · 计算机科学 2024-05-30 Shivesh Jadon , Mehrad Faridan , Edward Mah , Rajan Vaish , Wesley Willett , Ryo Suzuki

Non-verbal communication often comprises of semantically rich gestures that help convey the meaning of an utterance. Producing such semantic co-speech gestures has been a major challenge for the existing neural systems that can generate…

计算机视觉与模式识别 · 计算机科学 2025-04-07 M. Hamza Mughal , Rishabh Dabral , Merel C. J. Scholman , Vera Demberg , Christian Theobalt

Gaze-based virtual keyboards provide an effective interface for text entry by eye movements. The efficiency and usability of these keyboards have traditionally been evaluated with conventional text entry performance measures such as words…

人机交互 · 计算机科学 2018-04-10 Korok Sengupta , Jun Sun , Raphael Menges , Chandan Kumar , Steffen Staab

The rich textual information of large vision-language models (VLMs) combined with the powerful generative prior of pre-trained text-to-image (T2I) diffusion models has achieved impressive performance in single-image super-resolution (SISR).…

计算机视觉与模式识别 · 计算机科学 2025-08-25 Haodong He , Yancheng Bai , Rui Lan , Xu Duan , Lei Sun , Xiangxiang Chu , Gui-Song Xia

Vision based text entry systems aim to help disabled people achieve text communication using eye movement. Most previous methods have employed an existing eye tracker to predict gaze direction and design an input method based upon that.…

计算机视觉与模式识别 · 计算机科学 2017-07-04 Chi Zhang , Rui Yao , Jinpeng Cai

Estimating 3D hand pose from monocular RGB images is fundamental for applications in AR/VR, human-computer interaction, and sign language understanding. In this work we focus on a scenario where a discrete set of gesture labels is available…

计算机视觉与模式识别 · 计算机科学 2026-04-08 Rui Hong , Jana Kosecka

This paper presents a novel framework for speech-driven gesture production, applicable to virtual agents to enhance human-computer interaction. Specifically, we extend recent deep-learning-based, data-driven methods for speech-driven…

计算机视觉与模式识别 · 计算机科学 2021-04-07 Taras Kucherenko , Dai Hasegawa , Naoshi Kaneko , Gustav Eje Henter , Hedvig Kjellström