中文
相关论文

相关论文: ImageAssist: Tools for Enhancing Touchscreen-Based…

200 篇论文

Many blind and low vision (BLV) people are excluded from professional roles that may involve visual tasks due to access barriers and persisting stigmas. Advancing generative AI systems can support BLV people through providing contextual and…

人机交互 · 计算机科学 2025-10-13 Lucy Jiang , Lotus Zhang , Leah Findlater

In this paper, we propose a deep learning based assistive system to improve the environment perception experience of visually impaired (VI). The system is composed of a wearable terminal equipped with an RGBD camera and an earphone, a…

机器人学 · 计算机科学 2019-08-12 Yimin Lin , Kai Wang , Wanxin Yi , Shiguo Lian

Blind and low-vision (BLV) people rely on GPS-based systems for outdoor navigation. GPS's inaccuracy, however, causes them to veer off track, run into obstacles, and struggle to reach precise destinations. While prior work has made precise…

It is a challenging task for visually impaired people to perceive their surrounding environment due to the complexity of the natural scenes. Their personal and social activities are thus highly limited. This paper introduces a Large…

计算机视觉与模式识别 · 计算机科学 2025-04-28 Zezhou Chen , Zhaoxiang Liu , Kai Wang , Kohou Wang , Shiguo Lian

An innumerable number of individual choices go into discovering a new book. There are unmistakably two groups of booklovers: those who like to search online, follow other people's latest readings, or simply react to a system's…

人机交互 · 计算机科学 2020-11-03 Zona Kostic , Jared Jessup , Jeffrey Baglioni , Nathan Weeks , Johann Philipp Dreessen , Ning Chen , Tianyu Liu

Low-light image enhancement is challenging in that it needs to consider not only brightness recovery but also complex issues like color distortion and noise, which usually hide in the dark. Simply adjusting the brightness of a low-light…

图像与视频处理 · 电气工程与系统科学 2020-03-17 Feifan Lv , Yu Li , Feng Lu

People with Blind Visual Impairments (BVI) face unique challenges when sharing images, as these may accidentally contain sensitive or inappropriate content. In many instances, they are unaware of the potential risks associated with sharing…

人机交互 · 计算机科学 2026-03-05 Satabdi Das , Nahian Beente Firuj , Manjot Singh , Arshad Nasser , Khalad Hasan

Visually impaired people are often confronted with new environments and they find themselves face to face with an innumerous amount of difficulties when facing these environments. Having to surpass and deal with these difficulties that…

人机交互 · 计算机科学 2014-02-07 Ivo Rafael

Our previous interview study explores the needs and uses of diagrammatic information by the Blind and Low Vision (BLV) community, resulting in a framework called the Ladder of Diagram Access. The framework outlines five levels of…

人机交互 · 计算机科学 2024-04-02 Yichun Zhao , Miguel A. Nacenta

Over the past decade, considerable research has investigated Vision-Based Assistive Technologies (VBAT) to support people with vision impairments to understand and interact with their immediate environment using machine learning, computer…

A typical problem in Visual Analytics is that users are highly trained experts in their application domains, but have mostly no experience in using VA systems. Thus, users often have difficulties interpreting and working with visual…

Large multimodal models (LMMs) have enabled new AI-powered applications that help people with visual impairments (PVI) receive natural language descriptions of their surroundings through audible text. We investigated how this emerging…

人机交互 · 计算机科学 2025-02-25 Jingyi Xie , Rui Yu , He Zhang , Syed Masum Billah , Sooyeon Lee , John M. Carroll

In this paper, we propose an autonomous information seeking visual question answering framework, AVIS. Our method leverages a Large Language Model (LLM) to dynamically strategize the utilization of external tools and to investigate their…

计算机视觉与模式识别 · 计算机科学 2023-11-03 Ziniu Hu , Ahmet Iscen , Chen Sun , Kai-Wei Chang , Yizhou Sun , David A Ross , Cordelia Schmid , Alireza Fathi

An intuitive way to search for images is to use queries composed of an example image and a complementary text. While the first provides rich and implicit context for the search, the latter explicitly calls for new traits, or specifies how…

计算机视觉与模式识别 · 计算机科学 2022-05-17 Ginger Delmas , Rafael Sampaio de Rezende , Gabriela Csurka , Diane Larlus

People who are blind employ unique strategies when performing instrumental activities of daily living (iADLs), often relying on multiple sensory modalities and assistive technologies. While prior research has extensively explored adaptive…

人机交互 · 计算机科学 2025-02-25 Lily M. Turkstra , Tanya Bhatia , Alexa Van Os , Michael Beyeler

Low-light images are commonly encountered in real-world scenarios, and numerous low-light image enhancement (LLIE) methods have been proposed to improve the visibility of these images. The primary goal of LLIE is to generate clearer images…

计算机视觉与模式识别 · 计算机科学 2024-09-24 Xu Wu , Zhihui Lai , Zhou Jie , Can Gao , Xianxu Hou , Ya-nan Zhang , Linlin Shen

The increasing integration of artificial intelligence (AI) in visual analytics (VA) tools raises vital questions about the behavior of users, their trust, and the potential of induced biases when provided with guidance during data…

人机交互 · 计算机科学 2024-04-24 Sunwoo Ha , Shayan Monadjemi , Alvitta Ottley

Sighted and blind and low vision (BLV) creators alike use videos to communicate with broad audiences. Yet, video editing remains inaccessible to BLV creators. Our formative study revealed that current video editing tools make it difficult…

人机交互 · 计算机科学 2023-03-01 Mina Huh , Saelyne Yang , Yi-Hao Peng , Xiang 'Anthony' Chen , Young-Ho Kim , Amy Pavel

Perception-based image analysis technologies can be used to help visually impaired people take better quality pictures by providing automated guidance, thereby empowering them to interact more confidently on social media. The photographs…

计算机视觉与模式识别 · 计算机科学 2023-07-26 Maniratnam Mandal , Deepti Ghadiyaram , Danna Gurari , Alan C. Bovik

Recent VLM-based agents aim to replicate OpenAI O3's "thinking with images" via tool use, yet most open-source methods restrict inputs to a single image, limiting their applicability to real-world multi-image QA tasks. To address this gap,…

计算机视觉与模式识别 · 计算机科学 2026-04-06 Chengqi Dong , Chuhuai Yue , Hang He , Rongge Mao , Fenghe Tang , S Kevin Zhou , Zekun Xu , Xiaohan Wang , Jiajun Chai , Guojun Yin