中文
相关论文

相关论文: Revamp: Enhancing Accessible Information Seeking E…

200 篇论文

As online content becomes ever more visual, the demand for searching by visual queries grows correspondingly stronger. Shop The Look is an online shopping discovery service at Pinterest, leveraging visual search to enable users to find and…

计算机视觉与模式识别 · 计算机科学 2020-06-22 Raymond Shiau , Hao-Yu Wu , Eric Kim , Yue Li Du , Anqi Guo , Zhiyuan Zhang , Eileen Li , Kunlong Gu , Charles Rosenberg , Andrew Zhai

The widespread use of online review sites over the past decade has motivated businesses of all types to possess an expansive arsenal of user feedback to mark their reputation. Though a significant proportion of purchasing decisions are…

社会与信息网络 · 计算机科学 2016-02-24 Azade Nazi , Mahashweta Das , Gautam Das

Refreshable tactile displays (RTDs) are predicted to soon become a viable option for the provision of accessible graphics for people who are blind or have low vision (BLV). This new technology for the tactile display of braille and…

Individuals with visual impairments, encompassing both partial and total difficulties in visual perception, are referred to as visually impaired (VI) people. An estimated 2.2 billion individuals worldwide are affected by visual impairments.…

计算机视觉与模式识别 · 计算机科学 2024-04-04 Bufang Yang , Lixing He , Kaiwei Liu , Zhenyu Yan

Despite the recent surge of research efforts to make data visualizations accessible to people who are blind or have low vision (BLV), how to support BLV people's data analysis remains an important and challenging question. As refreshable…

人机交互 · 计算机科学 2025-01-17 Samuel Reinders , Matthew Butler , Ingrid Zukerman , Bongshin Lee , Lizhen Qu , Kim Marriott

With the development of Information and Communication Technologies and the dissemination of smartphones, especially now that image search is possible through the internet, e-commerce markets are more activating purchasing services for a…

信息检索 · 计算机科学 2021-11-16 Yonghyun Kim

People with blindness and low vision (pBLV) face significant challenges, struggling to navigate environments and locate objects due to limited visual cues. Spatial reasoning is crucial for these individuals, as it enables them to understand…

计算机视觉与模式识别 · 计算机科学 2025-05-19 Alexey Magay , Dhurba Tripathi , Yu Hao , Yi Fang

Engaging in recreational activities in public spaces poses challenges for blind people, often involving dependency on sighted help. Window shopping is a key recreational activity that remains inaccessible. In this paper, we investigate the…

人机交互 · 计算机科学 2024-05-13 Rie Kamikubo , Hernisa Kacorri , Chieko Asakawa

As social virtual reality (VR) grows more popular, addressing accessibility for blind and low vision (BLV) users is increasingly critical. Researchers have proposed an AI "sighted guide" to help users navigate VR and answer their questions,…

人机交互 · 计算机科学 2026-03-31 Jazmin Collins , Sharon Y Lin , Tianqi Liu , Andrea Stevenson Won , Shiri Azenkot

SightGlow is a web extension tailored to improve color perception accuracy for individuals with red-green color blindness. The research was focused on evaluating whether personalized color adjustment and selective zoom enhance user…

人机交互 · 计算机科学 2024-12-17 Sansrit Paudel

Visual reasoning, a cornerstone of human intelligence, encompasses complex perceptual and logical processes essential for solving diverse visual problems. While advances in computer vision have produced powerful models for various…

计算机视觉与模式识别 · 计算机科学 2025-09-03 Zetong Zhou , Dongping Chen , Zixian Ma , Zhihan Hu , Mingyang Fu , Sinan Wang , Yao Wan , Zhou Zhao , Ranjay Krishna

Navigating unfamiliar environments remains one of the most persistent and critical challenges for people who are blind or have limited vision (BLV). Existing assistive tools often rely on online services or APIs, making them costly,…

人机交互 · 计算机科学 2025-10-28 Dabbrata Das , Argho Deb Das , Farhan Sadaf , Azhar Uddin , Tirtho Mondal

E-commerce product understanding demands by nature, strong multimodal comprehension from text, images, and structured attributes. General-purpose Vision-Language Models (VLMs) enable generalizable multimodal latent modelling, yet there is…

Vision language models can now generate long-form answers to questions about images - long-form visual question answers (LFVQA). We contribute VizWiz-LF, a dataset of long-form answers to visual questions posed by blind and low vision (BLV)…

计算与语言 · 计算机科学 2025-07-28 Mina Huh , Fangyuan Xu , Yi-Hao Peng , Chongyan Chen , Hansika Murugu , Danna Gurari , Eunsol Choi , Amy Pavel

Navigation in new or unknown environments is vital, especially for visually impaired individuals. While many solutions exist, few are tailored to specific disabilities, often due to limited collaboration with handicap users in the design…

The proliferation of mobile devices and social media has revolutionized content dissemination, with short-form video becoming increasingly prevalent. This shift has introduced the challenge of video reframing to fit various screen aspect…

计算机视觉与模式识别 · 计算机科学 2024-03-12 Jiawang Cao , Yongliang Wu , Weiheng Chi , Wenbo Zhu , Ziyue Su , Jay Wu

Immersive Computer Graphics (CGs) rendering has become ubiquitous in modern daily life. However, comprehensively evaluating CG quality remains challenging for two reasons: First, existing CG datasets lack systematic descriptions of…

计算机视觉与模式识别 · 计算机科学 2026-03-12 Zhuangzi Li , Jian Jin , Shilv Cai , Weisi Lin

Understanding and managing data privacy in the digital world can be challenging for sighted users, let alone blind and low-vision (BLV) users. There is limited research on how BLV users, who have special accessibility needs, navigate data…

人机交互 · 计算机科学 2023-10-16 Yuanyuan Feng , Abhilasha Ravichander , Yaxing Yao , Shikun Zhang , Rex Chen , Shomir Wilson , Norman Sadeh

The Visual Object Information Retrieval (VOIR) system described in this paper implements an image retrieval approach that combines two layers, the conceptual and the visual layer. It uses terms from a textual thesaurus to represent the…

信息检索 · 计算机科学 2008-09-30 Jose Torres , Luis Paulo Reis

Efficiently learning visual representations of items is vital for large-scale recommendations. In this article we compare several pretrained efficient backbone architectures, both in the convolutional neural network (CNN) and in the vision…

计算机视觉与模式识别 · 计算机科学 2023-08-03 Eden Dolev , Alaa Awad , Denisa Roberts , Zahra Ebrahimzadeh , Marcin Mejran , Vaibhav Malpani , Mahir Yavuz