中文
相关论文

相关论文: Identifying Crucial Objects in Blind and Low-Visio…

200 篇论文

Recent studies have focused on facilitating perception and outdoor navigation for people with blindness or some form of vision loss. However, a significant portion of these studies is centered around treatment and vision rehabilitation,…

人机交互 · 计算机科学 2022-12-22 Paniz Sedighi , Mohammad Hesam Norouzi , Mehdi Delrobaei

The understanding of human-object interactions is fundamental in First Person Vision (FPV). Visual tracking algorithms which follow the objects manipulated by the camera wearer can provide useful information to effectively model such…

计算机视觉与模式识别 · 计算机科学 2022-10-20 Matteo Dunnhofer , Antonino Furnari , Giovanni Maria Farinella , Christian Micheloni

Object detection is a fundamental visual recognition problem in computer vision and has been widely studied in the past decades. Visual object detection aims to find objects of certain target classes with precise localization in a given…

计算机视觉与模式识别 · 计算机科学 2019-08-13 Xiongwei Wu , Doyen Sahoo , Steven C. H. Hoi

In this paper, we address an issue that the visually impaired commonly face while crossing intersections and propose a solution that takes form as a mobile application. The application utilizes a deep learning convolutional neural network…

计算机视觉与模式识别 · 计算机科学 2022-06-13 Samuel Yu , Heon Lee , Jung Hoon Kim

Safe navigation for the visually impaired individuals remains a critical challenge, especially concerning head-level obstacles, which traditional mobility aids often fail to detect. We introduce GuideTouch, a compact, affordable, standalone…

People with visual impairments (PwVI) often have difficulties navigating through unfamiliar indoor environments. However, current wayfinding tools are fairly limited. In this short paper, we present our in-progress work on a wayfinding…

Most existing assistive navigation tools focus on providing real-time guidance for Blind and Low-Vision (BLV) people, but few support building a holistic spatial understanding of unfamiliar environments before travel. Such cognitive map…

人机交互 · 计算机科学 2026-04-17 Li Liu , Jiaming Qu , Marc Jowell Bagaoisan , David T. Lee , Leilani H. Gilpin

Purpose: Navigating urban environments poses significant challenges for individuals who are blind or have low vision, especially in areas affected by construction. Construction zones introduce hazards such as uneven surfaces, barriers,…

计算机视觉与模式识别 · 计算机科学 2025-09-25 Junchi Feng , Giles Hamilton-Fletcher , Nikhil Ballem , Michael Batavia , Yifei Wang , Jiuling Zhong , Maurizio Porfiri , John-Ross Rizzo

The navigation of indoor spaces poses difficult challenges for individuals with visual impairments, as it requires processing of sensory information, dealing with uncertainties, and relying on assistance. To tackle these challenges, we…

计算机与社会 · 计算机科学 2026-02-17 Johannes Wortmann , Bernd Schäufele , Konstantin Klipp , Ilja Radusch , Katharina Blaß , Thomas Jung

Blind and low-vision (BLV) people face many challenges when venturing into public environments, often wishing it were easier to get help from people nearby. Ironically, while many sighted individuals are willing to help, such interactions…

Indoor navigation is challenging due to the absence of satellite positioning. This challenge is manifold greater for Visually Impaired People (VIPs) who lack the ability to get information from wayfinding signage. Other sensor signals…

计算机视觉与模式识别 · 计算机科学 2024-10-25 Jun Yu , Yifan Zhang , Badrinadh Aila , Vinod Namboodiri

The proposed system aims to be a techno-friend of visually impaired people to assist them in orientation and mobility both indoor and outdoor. Moving through an unknown environment becomes a real challenge for most of them, although they…

人机交互 · 计算机科学 2021-12-03 Shyama Kumari Arunachalam , Roopa V , Meena H B , Vijayalakshmi , T Malavika

We present an approach to labeling short video clips with English verbs as event descriptions. A key distinguishing aspect of this work is that it labels videos with verbs that describe the spatiotemporal interaction between event…

Vision-and-language navigation (VLN) stands as a key research problem of Embodied AI, aiming at enabling agents to navigate in unseen environments following linguistic instructions. In this field, generalization is a long-standing…

计算机视觉与模式识别 · 计算机科学 2024-07-02 Jiazhao Zhang , Kunyu Wang , Rongtao Xu , Gengze Zhou , Yicong Hong , Xiaomeng Fang , Qi Wu , Zhizheng Zhang , He Wang

Humans use context and scene knowledge to easily localize moving objects in conditions of complex illumination changes, scene clutter and occlusions. In this paper, we present a method to leverage human knowledge in the form of annotated…

计算机视觉与模式识别 · 计算机科学 2016-04-20 Archith J. Bency , S. Karthikeyan , Carter De Leo , Santhoshkumar Sunderrajan , B. S. Manjunath

Current state-of-the-art object detection and segmentation methods work well under the closed-world assumption. This closed-world setting assumes that the list of object categories is available during training and deployment. However, many…

计算机视觉与模式识别 · 计算机科学 2021-04-13 Weiyao Wang , Matt Feiszli , Heng Wang , Du Tran

Prior research has shown how 'content preview tools' improve speed and accuracy of user relevance judgements across different information retrieval tasks. This paper describes a novel user interface tool, the Content Flow Bar, designed to…

Searching for objects in unfamiliar scenarios is a challenging task for blind people. It involves specifying the target object, detecting it, and then gathering detailed information according to the user's intent. However, existing…

Object locating in virtual reality (VR) has been widely used in many VR applications, such as virtual assembly, virtual repair, virtual remote coaching. However, when there are a large number of objects in the virtual environment(VE), the…

人机交互 · 计算机科学 2022-12-08 Xiaoheng Wei , Xuehuai Shi , Lili Wang

Multimodal large language models (MLLMs) provide new opportunities for blind and low vision (BLV) people to access visual information in their daily lives. However, these models often produce errors that are difficult to detect without…

人机交互 · 计算机科学 2025-07-22 Meng Chen , Akhil Iyer , Amy Pavel