中文
相关论文

相关论文: Computer Vision and Conflicting Values: Describing…

200 篇论文

Many blind and low vision (BLV) people are excluded from professional roles that may involve visual tasks due to access barriers and persisting stigmas. Advancing generative AI systems can support BLV people through providing contextual and…

人机交互 · 计算机科学 2025-10-13 Lucy Jiang , Lotus Zhang , Leah Findlater

Like sighted people, visually impaired people want to share photographs on social networking services, but find it difficult to identify and select photos from their albums. We aimed to address this problem by incorporating state-of-the-art…

人机交互 · 计算机科学 2018-05-07 Yuhang Zhao , Shaomei Wu , Lindsay Reynolds , Shiri Azenkot

Illustrations are widely used in education, and sometimes, alternatives are not available for visually impaired students. Therefore, those students would benefit greatly from an automatic illustration description system, but only if those…

人机交互 · 计算机科学 2021-06-30 Anett Hoppe , David Morris , Ralph Ewerth

Alternative Texts (Alt-Text) for chart images are essential for making graphics accessible to people with blindness and visual impairments. Traditionally, Alt-Text is manually written by authors but often encounters issues such as…

计算机视觉与模式识别 · 计算机科学 2024-05-30 Omar Moured , Shahid Ali Farooqui , Karin Muller , Sharifeh Fadaeijouybari , Thorsten Schwarz , Mohammed Javed , Rainer Stiefelhagen

Increasing computational power and improving deep learning methods have made computer vision technologies pervasively common in urban environments. Their applications in policing, traffic management, and documenting public spaces are…

计算机与社会 · 计算机科学 2023-01-06 Anthony Vanky , Ri Le

Automated computer vision systems have been applied in many domains including security, law enforcement, and personal devices, but recent reports suggest that these systems may produce biased results, discriminating against people in…

计算机视觉与模式识别 · 计算机科学 2020-05-22 Jungseock Joo , Kimmo Kärkkäinen

Automatic description generation from natural images is a challenging problem that has recently received a large amount of interest from the computer vision and natural language processing communities. In this survey, we classify the…

Disability Services Office (DSO) professionals at higher education institutions write alt text for {visual content}. However, due to the complexity of visual content, such as HCI figures in research publications, DSO professionals can…

人机交互 · 计算机科学 2026-02-10 Muhammad Raees , Yugo Iwamoto , Konstantinos Papangelis , Jamison Heard , Garreth W. Tigwell

This paper aims to shed light on the ethical problems of creating and deploying computer vision tech, particularly in using publicly available datasets. Due to the rapid growth of machine learning and artificial intelligence, computer…

计算机视觉与模式识别 · 计算机科学 2025-08-12 Ghalib Ahmed Tahir

Automatically generating a human-like description for a given image is a potential research in artificial intelligence, which has attracted a great of attention recently. Most of the existing attention methods explore the mapping…

计算机视觉与模式识别 · 计算机科学 2020-11-03 Feicheng Huang , Zhixin Li , Haiyang Wei , Canlong Zhang , Huifang Ma

"Scene description" applications that describe visual content in a photo are useful daily tools for blind and low vision (BLV) people. Researchers have studied their use, but they have only explored those that leverage remote sighted…

人机交互 · 计算机科学 2025-03-13 Ricardo Gonzalez , Jazmin Collins , Shiri Azenkot , Cynthia Bennett

Figures in scientific publications contain important information and results, and alt text is needed for blind and low vision readers to engage with their content. We conduct a study to characterize the semantic content of alt text in HCI…

人机交互 · 计算机科学 2022-09-29 Sanjana Chintalapati , Jonathan Bragg , Lucy Lu Wang

Social media platforms today strive to improve user experience through AI recommendations, yet the value of such recommendations vanishes as users do not understand the reasons behind them. This issue arises because explainability in social…

人工智能 · 计算机科学 2025-08-04 Banan Alkhateeb , Ellis Solaiman

The contextual information of Web images is investigated to address the issue of characterizing their content with semantic descriptors and therefore bridge the semantic gap, i.e. the gap between their automated low-level representation in…

信息检索 · 计算机科学 2020-05-06 Fariza Fauzi , Mohammed Belkhatir

People with visual impairments often struggle to create content that relies heavily on visual elements, particularly when conveying spatial and structural information. Existing accessible drawing tools, which construct images line by line,…

人机交互 · 计算机科学 2025-04-30 Seonghee Lee , Maho Kohga , Steve Landau , Sile O'Modhrain , Hari Subramonyam

Alternative text is critical in communicating graphics to people who are blind or have low vision. Especially for graphics that contain rich information, such as visualizations, poorly written or an absence of alternative texts can worsen…

人机交互 · 计算机科学 2021-08-10 Crescentia Jung , Shubham Mehta , Atharva Kulkarni , Yuhang Zhao , Yea-Seul Kim

Computer vision based technology is becoming ubiquitous in society. One application area that has seen an increase in computer vision is assistive technologies, specifically for those with visual impairment. Research has shown the ability…

计算机视觉与模式识别 · 计算机科学 2019-05-21 Linda Wang , Alexander Wong

The web is littered with images, once created for human consumption and now increasingly interpreted by agents using vision-language models (VLMs). These agents make visual decisions at scale, deciding what to click, recommend, or buy. Yet,…

计算机视觉与模式识别 · 计算机科学 2026-02-18 Manuel Cherep , Pranav M R , Pattie Maes , Nikhil Singh

Large vision-language models (VLMs) can assist visually impaired people by describing images from their daily lives. Current evaluation datasets may not reflect diverse cultural user backgrounds or the situational context of this use case.…

计算机视觉与模式识别 · 计算机科学 2024-07-09 Antonia Karamolegkou , Phillip Rust , Yong Cao , Ruixiang Cui , Anders Søgaard , Daniel Hershcovich

Blind and low vision (BLV) internet users access images on the web via text descriptions. New vision-to-language models such as GPT-V, Gemini, and LLaVa can now provide detailed image descriptions on-demand. While prior research and…

人机交互 · 计算机科学 2024-09-06 Ananya Gubbi Mohanbabu , Amy Pavel
‹ 上一页 1 2 3 10 下一页 ›