中文
相关论文

相关论文: Exploring Community-Driven Descriptions for Making…

200 篇论文

We consider the task of identifying human actions visible in online videos. We focus on the widely spread genre of lifestyle vlogs, which consist of videos of people performing actions while verbally describing them. Our goal is to identify…

计算与语言 · 计算机科学 2021-09-10 Oana Ignat , Laura Burdick , Jia Deng , Rada Mihalcea

Access to textual and visual information for visually impaired persons becomes very difficult with screen readers which are not adapted to different websites.This paper analyses the use of different technologies for access digital content…

人机交互 · 计算机科学 2019-11-18 Katerine Romeo , Edwige Pissaloux , Frédéric Serin

Accessibility of tables on websites for Visually Impaired Persons (VIP) is not optimal with screen readers which are not always effective for the recovery of visual information (2D). Actual Web/Multimedia technologies are not taking in…

人机交互 · 计算机科学 2019-11-11 Katerine Romeo , E Pissaloux , F Serin

Accessibility of websites for visually impaired persons is mishandled by screen readers which are not always adapted to interactivity needed by actual web/multimedia technologies. This paper analyses the difficulties to access to…

人机交互 · 计算机科学 2019-12-09 Katerine Romeo , Edwige Pissaloux , Frédéric Serin

Storyline visualization has emerged as an innovative method for illustrating the development and changes in stories across various domains. Traditional approaches typically represent stories with one line per character, progressing from…

人机交互 · 计算机科学 2024-08-06 Haonan Yao , Lixiang Zhao , Boyuan Chen , Kaiwen Li , Hai-Ning Liang , Lingyun Yu

Various programming tools, languages, and environments give programmers the impression of changing a program while it is running. This experience of liveness has been discussed for over two decades and a broad spectrum of research on this…

编程语言 · 计算机科学 2018-08-01 Patrick Rein , Stefan Ramson , Jens Lincke , Robert Hirschfeld , Tobias Pape

This paper proposes a method for generating bullet comments for live-streaming games based on highlights (i.e., the exciting parts of video clips) extracted from the game content and evaluate the effect of mental health promotion. Game live…

多媒体 · 计算机科学 2021-08-19 Junjie H. Xu , Yulin Cai , Zhou Fang , Pujana Paliyawan

Digital avatars are an important part of identity representation, but there is little work on understanding how to represent disability. We interviewed 18 people with disabilities and related identities about their experiences and…

人机交互 · 计算机科学 2023-02-06 Kelly Mack , Rai Ching Ling Hsu , Andrés Monroy-Hernández , Brian A. Smith , Fannie Liu

Blind and low vision (BLV) internet users access images on the web via text descriptions. New vision-to-language models such as GPT-V, Gemini, and LLaVa can now provide detailed image descriptions on-demand. While prior research and…

人机交互 · 计算机科学 2024-09-06 Ananya Gubbi Mohanbabu , Amy Pavel

In recent days, streaming technology has greatly promoted the development in the field of livestream. Due to the excessive length of livestream records, it's quite essential to extract highlight segments with the aim of effective…

多媒体 · 计算机科学 2022-06-13 Yang Zhao , Xuan Lin , Wenqiang Xu , Maozong Zheng , Zhengyong Liu , Zhou Zhao

It has always been a rather tough task to communicate with someone possessing a hearing impairment. One of the most tested ways to establish such a communication is through the use of sign based languages. However, not many people are aware…

计算机视觉与模式识别 · 计算机科学 2025-07-22 Sharanya Mukherjee , Md Hishaam Akhtar , Kannadasan R

The main objective of this article is to provide an overview of P2P based Video-on-Demand and live streaming services. The article starts with an introduction to media streaming and its simplified architecture. Various solutions offering…

网络与互联网体系结构 · 计算机科学 2013-04-05 Sabu M Thampi

Few VR applications and games implement captioning of speech and audio cues, which either inhibits or prevents access of their application by deaf or hard of hearing (DHH) users, new language learners, and other caption users. Additionally,…

人机交互 · 计算机科学 2022-10-28 Pranav Pidathala , Dawson Franz , James Waller , Raja Kushalnagar , Christian Vogler

Despite advancements in Video Large Language Models (Vid-LLMs) improving multimodal understanding, challenges persist in streaming video reasoning due to its reliance on contextual information. Existing paradigms feed all available…

计算机视觉与模式识别 · 计算机科学 2025-12-30 Zicheng Zhao , Kangyu Wang , Shijie Li , Rui Qian , Weiyao Lin , Huabin Liu

With the rapid growth of live streaming platforms, personalized recommendation systems have become pivotal in improving user experience and driving platform revenue. The dynamic and multimodal nature of live streaming content (e.g., visual,…

信息检索 · 计算机科学 2025-08-22 Yalong Guan , Xiang Chen , Mingyang Wang , Xiangyu Wu , Lihao Liu , Chao Qi , Shuang Yang , Tingting Gao , Guorui Zhou , Changjian Chen

Billions of live-streaming viewers share their opinions on scenes they are watching in real-time and interact with the event, commentators as well as other viewers via text comments. Thus, there is necessary to explore viewers' comments…

多媒体 · 计算机科学 2023-01-18 Junjie H. Xu , Yu Nakano , Lingrong Kong , Kojiro Iizuka

Thanks to mobile apps such as Periscope and Facebook Live, live-streaming video is having a moment again. It has not been clear, however, to what extent the current ubiquity of smartphones is impacting this technology's acceptance in…

人机交互 · 计算机科学 2019-02-19 Cori Faklaris , Asa Blevins , Matthew O'Haver , Neha Singhal , Francesco Cafaro

We describe here our perception of complex systems, of how we feel the different layers of description are important part of a correct complex system simulation. We describe a rough models categorization between rules based and law based,…

流体动力学 · 物理学 2007-12-18 Pierrick Tranouez , Cyrille Bertelle , Damien Olivier

Recent advances in Large Multi-modal Models (LMMs) are primarily focused on offline video understanding. Instead, streaming video understanding poses great challenges to recent models due to its time-sensitive, omni-modal and interactive…

计算机视觉与模式识别 · 计算机科学 2025-03-18 Shenghao Fu , Qize Yang , Yuan-Ming Li , Yi-Xing Peng , Kun-Yu Lin , Xihan Wei , Jian-Fang Hu , Xiaohua Xie , Wei-Shi Zheng

We present Hanstreamer, a free and open-source system for webcam-based data presentation. The system performs real-time gesture recognition on the user's webcam video stream to provide interactive data visuals. Apart from the standard chart…

人机交互 · 计算机科学 2023-09-25 Adrian Kristanto , Maxime Cordeil , Benjamin Tag , Nathalie Henry Riche , Tim Dwyer