English
Related papers

Related papers: Audo-Sight: Enabling Ambient Interaction For Blind…

200 papers

Finding obstacle-free paths in unknown environments is a big navigation issue for visually impaired people and autonomous robots. Previous works focus on obstacle avoidance, however they do not have a general view of the environment they…

This article presents an extensive literature review of technology based intervention methodologies for individuals facing Autism Spectrum Disorder (ASD). Reviewed methodologies include: contemporary Computer Aided Systems (CAS), Computer…

Human-Computer Interaction · Computer Science 2019-02-21 Muhammad Shoaib Jaliawala , Rizwan Ahmed Khan

Vision-language models (VLMs) have recently emerged as powerful representation learning systems that align visual observations with natural language concepts, offering new opportunities for semantic reasoning in safety-critical autonomous…

Computer Vision and Pattern Recognition · Computer Science 2026-02-19 Ross Greer , Maitrayee Keskar , Angel Martinez-Sanchez , Parthib Roy , Shashank Shriram , Mohan Trivedi

With recent advances in multi-modal foundation models, the previously text-only large language models (LLM) have evolved to incorporate visual input, opening up unprecedented opportunities for various applications in visualization. Our work…

Human-Computer Interaction · Computer Science 2023-12-08 Shusen Liu , Haichao Miao , Zhimin Li , Matthew Olson , Valerio Pascucci , Peer-Timo Bremer

When establishing a visual connection between a virtual reality user and an augmented reality user, it is important to consider whether the augmented reality user faces a surplus of information. Augmented reality, compared to virtual…

Human-Computer Interaction · Computer Science 2021-05-05 Robbe Cools , Jihae Han , Adalberto L. Simeone

Blind people are often called to contribute image data to datasets for AI innovation with the hope for future accessibility and inclusion. Yet, the visual inspection of the contributed images is inaccessible. To this day, we lack mechanisms…

Human-Computer Interaction · Computer Science 2024-07-30 Rie Kamikubo , Farnaz Zamiri Zeraati , Kyungjun Lee , Hernisa Kacorri

To bridge the gap between vision and language modalities, Multimodal Large Language Models (MLLMs) usually learn an adapter that converts visual inputs to understandable tokens for Large Language Models (LLMs). However, most adapters…

Computer Vision and Pattern Recognition · Computer Science 2024-05-27 Yue Zhang , Hehe Fan , Yi Yang

Animist worldviews treat beings, plants, landscapes, and even tools as persons endowed with spirit, an orientation that has long shaped human-nonhuman relations through ritual and moral practice. While modern industrial societies have often…

Artificial Intelligence · Computer Science 2025-10-01 Diana Mykhaylychenko , Maisha Thasin , Dunya Baradari , Charmelle Mhungu

The advent of Large Language Models (LLMs) has significantly reshaped the trajectory of the AI revolution. Nevertheless, these LLMs exhibit a notable limitation, as they are primarily adept at processing textual information. To address this…

Computer Vision and Pattern Recognition · Computer Science 2025-10-15 Akash Ghosh , Arkadeep Acharya , Sriparna Saha , Vinija Jain , Aman Chadha

Visual feedback plays a crucial role in the process of amputation patients completing grasping in the field of prosthesis control. However, for blind and visually impaired (BVI) amputees, the loss of both visual and grasping abilities makes…

Robotics · Computer Science 2023-08-15 Chunhao Peng , Dapeng Yang , Ming Cheng , Jinghui Dai , Deyu Zhao , Li Jiang

Current Vision-Language-Action (VLA) models are often constrained by a rigid, static interaction paradigm, which lacks the ability to see, hear, speak, and act concurrently as well as handle real-time user interruptions dynamically. This…

This survey and application guide to multimodal large language models(MLLMs) explores the rapidly developing field of MLLMs, examining their architectures, applications, and impact on AI and Generative Models. Starting with foundational…

Artificial Intelligence · Computer Science 2025-12-02 Chia Xin Liang , Pu Tian , Caitlyn Heqi Yin , Yao Yua , Wei An-Hou , Li Ming , Xinyuan Song , Tianyang Wang , Ziqian Bi , Ming Liu

As digital worlds become ubiquitous via video games, simulations, virtual and augmented reality, people with disabilities who cannot access those worlds are becoming increasingly disenfranchised. More often than not the design of these…

Human-Computer Interaction · Computer Science 2020-01-23 Bijan Fakhri , Troy McDaniel , Heni Ben Amor , Hemanth Venkateswara , Abhik Chowdhury , Sethuraman Panchanathan

Embodied artificial intelligence (Embodied AI) plays a pivotal role in the application of advanced technologies in the intelligent era, where AI systems are integrated with physical bodies that enable them to perceive, reason, and interact…

Artificial Intelligence · Computer Science 2025-06-24 Zhaohan Feng , Ruiqi Xue , Lei Yuan , Yang Yu , Ning Ding , Meiqin Liu , Bingzhao Gao , Jian Sun , Xinhu Zheng , Gang Wang

Human drivers adeptly navigate complex scenarios by utilizing rich attentional semantics, but the current autonomous systems struggle to replicate this ability, as they often lose critical semantic information when converting 2D…

Computer Vision and Pattern Recognition · Computer Science 2025-09-19 Pei Liu , Haipeng Liu , Haichao Liu , Xin Liu , Jinxin Ni , Jun Ma

We introduce ArtInsight, a novel AI-powered system to facilitate deeper engagement with child-created artwork in mixed visual-ability families. ArtInsight leverages large language models (LLMs) to craft a respectful and thorough initial…

Human-Computer Interaction · Computer Science 2025-03-11 Arnavi Chheda-Kothary , Ritesh Kanchi , Chris Sanders , Kevin Xiao , Aditya Sengupta , Melanie Kneitmix , Jacob O. Wobbrock , Jon E. Froehlich

Video-based learning (VBL) has become a dominant method for learning practical skills, yet accessibility guidelines provide limited guidance for users with cognitive differences. In particular, challenges that individuals with Borderline…

Human-Computer Interaction · Computer Science 2026-02-10 Hyehyun Chu , Seungju Kim , Chen Zhou , Yu-Kai Hung , Saelyne Yang , Hyun W. Ka , Juho Kim

Interactive streetscape mapping tools such as Google Street View (GSV) and Meta Mapillary enable users to virtually navigate and experience real-world environments via immersive 360{\deg} imagery but remain fundamentally inaccessible to…

Human-Computer Interaction · Computer Science 2025-09-29 Jon E. Froehlich , Alexander Fiannaca , Nimer Jaber , Victor Tsaran , Shaun Kane

There is great promise in creating effective technology experiences during situationally-induced impairments and disabilities through the combination of universal design and adaptive interfaces. We believe this combination is a powerful…

Human-Computer Interaction · Computer Science 2019-04-15 Aaron Steinfeld , John Zimmerman , Anthony Tomasic

People with visual impairments face numerous challenges in their daily lives, including mobility, access to information, independent living, and employment. Artificial Intelligence (AI) with Computer Vision (CV) has the potential to improve…

Human-Computer Interaction · Computer Science 2024-07-26 Saidarshan Bhagat , Padmaja Joshi , Avinash Agarwal , Shubhanshu Gupta
‹ Prev 1 8 9 10 Next ›