English
Related papers

Related papers: Multimodal Appearance based Gaze-Controlled Virtua…

200 papers

The integration of Transparent Displays (TD) in various applications, such as Heads-Up Displays (HUDs) in vehicles, is a burgeoning field, poised to revolutionize user experiences. However, this innovation brings forth significant…

Human-Computer Interaction · Computer Science 2024-06-28 Esmaeil Seraj , Harsh Bhate , Walter Talamonti

Input multimodality combining speech and hand gestures has motivated numerous usability studies. Contrastingly, issues relating to the design and ergonomic evaluation of multimodal output messages combining speech with visual modalities…

Human-Computer Interaction · Computer Science 2007-09-05 Suzanne Kieffer , Noëlle Carbonell

The paper introduces Hands-Free VR, a voice-based natural-language interface for VR. The user gives a command using their voice, the speech audio data is converted to text using a speech-to-text deep learning model that is fine-tuned for…

Adaptive interfaces can help users perform sequential decision-making tasks like robotic teleoperation given noisy, high-dimensional command signals (e.g., from a brain-computer interface). Recent advances in human-in-the-loop machine…

Robotics · Computer Science 2023-09-08 Jensen Gao , Siddharth Reddy , Glen Berseth , Anca D. Dragan , Sergey Levine

We present our current research on the implementation of gaze as an efficient and usable pointing modality supplementary to speech, for interacting with augmented objects in our daily environment or large displays, especially immersive…

Human-Computer Interaction · Computer Science 2007-08-28 Daniel Gepner , Jérôme Simonin , Noëlle Carbonell

The control of multi-fingered dexterous prosthetics hand remains challenging due to the lack of an intuitive and efficient Grasp-type Switching Interface (GSI). We propose a new GSI (i-GSI) t hat integrates the manifold power of…

Robotics · Computer Science 2022-04-25 Chunyuan Shi , Dapeng Yang , Siyang Qiu , Jingdong Zhao

Automatic unknown word detection techniques can enable new applications for assisting English as a Second Language (ESL) learners, thus improving their reading experiences. However, most modern unknown word detection methods require…

Human-Computer Interaction · Computer Science 2023-03-21 Jiexin Ding , Bowen Zhao , Yuqi Huang , Yuntao Wang , Yuanchun Shi

Training vision-language models on cognitively-plausible amounts of data requires rethinking how models integrate multimodal information. Within the constraints of the Vision track for the BabyLM Challenge 2025, we propose a lightweight…

Artificial Intelligence · Computer Science 2025-10-10 Bianca-Mihaela Ganescu , Suchir Salhan , Andrew Caines , Paula Buttery

Visual Language Tracking (VLT) enhances tracking by mitigating the limitations of relying solely on the visual modality, utilizing high-level semantic information through language. This integration of the language enables more advanced…

Computer Vision and Pattern Recognition · Computer Science 2024-09-16 Xuchen Li , Shiyu Hu , Xiaokun Feng , Dailing Zhang , Meiqi Wu , Jing Zhang , Kaiqi Huang

Text input on mobile devices without physical keys can be challenging for people who are blind or low-vision. We interview 12 blind adults about their experiences with current mobile text input to provide insights into what sorts of…

Human-Computer Interaction · Computer Science 2025-10-03 Dylan Gaines , Keith Vertanen

Although eye-tracking technology is being integrated into more VR and MR headsets, the true potential of eye tracking in enhancing user interactions within XR settings remains relatively untapped. Presently, one of the most prevalent gaze…

Human-Computer Interaction · Computer Science 2024-05-24 Naveen Sendhilnathan , Ajoy S. Fernandes , Michael J. Proulx , Tanya R. Jonker

In order to avoid the "Midas Touch" problem, gaze-based interfaces for selection often introduce a dwell time: a fixed amount of time the user must fixate upon an object before it is selected. Past interfaces have used a uniform dwell time…

Human-Computer Interaction · Computer Science 2022-09-07 Zhaokang Chen , Bertram E. Shi

As virtual reality (VR) continues to evolve, traditional input methods such as handheld controllers and gesture systems often face challenges with precision, social accessibility, and user fatigue. These limitations motivate the exploration…

Human-Computer Interaction · Computer Science 2025-09-15 Xiang Li , Wei He , Per Ola Kristensson

Conducting collaborative tasks, e.g., multi-user game, in virtual reality (VR) could enable us to explore more immersive and effective experience. However, for current VR systems, users cannot communicate properly with each other via their…

Human-Computer Interaction · Computer Science 2023-03-21 Song Zhao , Shiwei Cheng , Chenshuang Zhu

Eye movements play a vital role in perceiving the world. Eye gaze can give a direct indication of the users point of attention, which can be useful in improving human-computer interaction. Gaze estimation in a non-intrusive manner can make…

Computer Vision and Pattern Recognition · Computer Science 2019-07-11 Anjith George

Gradual typing combines static and dynamic typing in the same program. One would hope that the performance in a gradually typed language would range between that of a dynamically typed language and a statically typed language. Existing…

Programming Languages · Computer Science 2018-02-20 Andre Kuhlenschmidt , Deyaaeldeen Almahallawi , Jeremy G. Siek

Augmented Reality (AR) enables intuitive interaction with virtual annotations overlaid on the real world, supporting a wide range of applications such as remote assistance, education, and industrial training. However, as the number of…

Human-Computer Interaction · Computer Science 2025-09-16 Zahra Borhani , Ali Ebrahimpour-Boroojeny , Francisco R. Ortega

In this study, we investigated gaze-based interaction methods within a virtual reality game with a visual search task with 52 participants. We compared four different interaction techniques: Selection by dwell time or confirmation of…

Human-Computer Interaction · Computer Science 2025-04-17 Björn Rene Severitt , Yannick Sauer , Alexander Neugebauer , Rajat Agarwala , Nora Castner , Siegfried Wahl

Interacting with multiple objects simultaneously makes us fast. A pre-step to this interaction is to select the objects, i.e., multi-object selection, which is enabled through two steps: (1) toggling multi-selection mode -- mode-switching…

Human-Computer Interaction · Computer Science 2026-02-16 Mohammad Raihanul Bashar , Aunnoy K Mutasim , Ken Pfeuffer , Anil Ufuk Batmaz

Recent work has shown that speech paired with images can be used to learn semantically meaningful speech representations even without any textual supervision. In real-world low-resource settings, however, we often have access to some…

Computation and Language · Computer Science 2019-09-04 Ankita Pasad , Bowen Shi , Herman Kamper , Karen Livescu
‹ Prev 1 4 5 6 7 8 10 Next ›