中文
相关论文

相关论文: Aikyam: A Video Conferencing Utility for Deaf and …

200 篇论文

Sign language visual recognition from continuous multi-modal streams is still one of the most challenging fields. Recent advances in human actions recognition are exploiting the ascension of GPU-based learning from massive data, and are…

计算机视觉与模式识别 · 计算机科学 2020-09-23 Bassem Seddik , Najoua Essoukri Ben Amara

Both text and video data are abundant on the internet and support large-scale self-supervised learning through next token or frame prediction. However, they have not been equally leveraged: language models have had significant real-world…

计算机视觉与模式识别 · 计算机科学 2024-02-28 Sherry Yang , Jacob Walker , Jack Parker-Holder , Yilun Du , Jake Bruce , Andre Barreto , Pieter Abbeel , Dale Schuurmans

The outbreak of COVID-19 forced schools to swiftly transition from in-person classes to online or remote offerings, making educators and learners alike rely on online videoconferencing platforms. Platforms like Zoom offer audio-visual…

人机交互 · 计算机科学 2023-01-23 Yanting Wu , Yuan Sun , S. Shyam Sundar

Many personal devices have transitioned from visual-controlled interfaces to speech-controlled interfaces to reduce device costs and interactive friction. This transition has been hastened by the increasing capabilities of speech-controlled…

人机交互 · 计算机科学 2019-09-04 Abraham Glasser , Kesavan Kushalnagar , Raja Kushalnagar

Multi-modal large language models have garnered significant interest recently. Though, most of the works focus on vision-language multi-modal models providing strong capabilities in following vision-and-language instructions. However, we…

计算与语言 · 计算机科学 2023-09-19 Yu Shu , Siwei Dong , Guangyao Chen , Wenhao Huang , Ruihua Zhang , Daochen Shi , Qiqi Xiang , Yemin Shi

Many companies have a suite of digital tools, such as Enterprise Social Networks, conferencing and document sharing software, and email, to facilitate collaboration among employees. During, or at the end of a collaboration, documents are…

人机交互 · 计算机科学 2019-05-17 Antoine Flepp , Julie Dugdale , Fabrice Bourge , Tiphaine Marie-Cardot

We present the Telepresence Lantern concept, developed to provide opportunities for older adults to stay in contact with remote family and friends. It provides a new approach to video-mediated communication, designed to facilitate natural…

人机交互 · 计算机科学 2023-08-31 Thomas H. Weisswange , Joel B. Schwartz , Aaron J. Horowitz , Jens Schmüdderich

A primary challenge for the deaf and hearing-impaired community stems from the communication gap with the hearing society, which can greatly impact their daily lives and result in social exclusion. To foster inclusivity in society, our…

计算机视觉与模式识别 · 计算机科学 2025-09-19 Elisa Cabana

Deafblind people have both hearing and visual impairments, which makes communication with other people often dependent on expensive technologies e.g., Braille displays, or on caregivers acting as interpreters. This paper presents Morse I/O…

人机交互 · 计算机科学 2022-05-11 David C. Kutner , Sunčica Hadžidedić

The main backbone of our Artificial Eye model is the Raspberry pi3 which is connected to the webcam ,ultrasonic proximity sensor, speaker and we also run all our software models i.e object detection, Optical Character recognition, google…

计算机视觉与模式识别 · 计算机科学 2023-08-03 Abhinav Benagi , Dhanyatha Narayan , Charith Rage , A Sushmitha

Researchers have adopted remote methods, such as online surveys and video conferencing, to overcome challenges in conducting in-person usability testing, such as participation, user representation, and safety. However, remote user…

人机交互 · 计算机科学 2022-03-09 Kyungjun Lee , Jonggi Hong , Ebrima Jarjue , Ernest Essuah Mensah , Hernisa Kacorri

During speech, people spontaneously gesticulate, which plays a key role in conveying information. Similarly, realistic co-speech gestures are crucial to enable natural and smooth interactions with social agents. Current end-to-end co-speech…

Smartphone use has grown rapidly, but the ways it shapes concurrent face-to-face interaction remains scarcely studied. In our research we have formulated two new concepts to depict this: 1) Sticky media device illustrates situations in…

人机交互 · 计算机科学 2019-10-30 Sanna Raudaskoski , Eerik Mantere , Satu Valkonen

This paper presents a new approach for end-to-end audio-visual multi-talker speech recognition. The approach, referred to here as the visual context attention model (VCAM), is important because it uses the available video information to…

声音 · 计算机科学 2022-04-05 Richard Rose , Olivier Siohan

The rate of disability is increase day by day all over the world .There are various type of Disabilities but the deaf persons are on second number among all types of disabilities.. In most of the countries disabled persons are supposed to…

计算机与社会 · 计算机科学 2013-10-22 Syed Asif Ali , Safeeulah Soomro , Abdul Ghafoor Memon , Mashooque Ahmed

Assistive technologies for the visually impaired have evolved to facilitate interaction with a complex and dynamic world. In this paper, we introduce AIris, an AI-powered wearable device that provides environmental awareness and interaction…

Deliberation is essential to well-functioning democracies, yet physical, economic, and social barriers often exclude certain groups, reducing representativeness and contributing to issues like group polarization. In this work, we explore…

人机交互 · 计算机科学 2025-11-17 Suyash Fulay , Dimitra Dimitrakopoulou , Deb Roy

Visual-to-auditory sensory substitution devices can assist the blind in sensing the visual environment by translating the visual information into a sound pattern. To improve the translation quality, the task performances of the blind are…

计算机视觉与模式识别 · 计算机科学 2019-04-22 Di Hu , Dong Wang , Xuelong Li , Feiping Nie , Qi Wang

An important part of second language learning is conversation which is best practised with speakers whose native language is the language being learned. We facilitate this by pairing students from different countries learning each others'…

人机交互 · 计算机科学 2021-06-28 Aparajita Dey-Plissonneau , Hyowon Lee , Vincent Pradier , Michael Scriney , Alan F. Smeaton

Since American Sign Language (ASL) has no standard written form, Deaf signers frequently share videos in order to communicate in their native language. However, since both hands and face convey critical linguistic information in signed…

计算机视觉与模式识别 · 计算机科学 2023-11-28 Zhaoyang Xia , Carol Neidle , Dimitris N. Metaxas