中文
相关论文

相关论文: Understood: Real-Time Communication Support for Ad…

200 篇论文

The augmented-reality head-mounted display (e.g., Microsoft HoloLens) is one of the most innovative technologies in multimedia and human-computer interaction in recent years. Despite the emerging research of its applications on engineering,…

人机交互 · 计算机科学 2019-06-04 Yunlong Wang , Harald Reiterer

Resting-state fMRI is commonly used for diagnosing Autism Spectrum Disorder (ASD) by using network-based functional connectivity. It has been shown that ASD is associated with brain regions and their inter-connections. However,…

神经元与认知 · 定量生物学 2022-01-04 Ranjeet Ranjan Jha , Abhishek Bhardwaj , Devin Garg , Arnav Bhavsar , Aditya Nigam

Understanding social interaction in video requires reasoning over a dynamic interplay of verbal and non-verbal cues: who is speaking, to whom, and with what gaze or gestures. While Multimodal Large Language Models (MLLMs) are natural…

计算机视觉与模式识别 · 计算机科学 2025-11-25 Liangyang Ouyang , Yifei Huang , Mingfang Zhang , Caixin Kang , Ryosuke Furuta , Yoichi Sato

Mixed Reality is increasingly used in mobile settings beyond controlled home and office spaces. This mobility introduces the need for user interface layouts that adapt to varying contexts. However, existing adaptive systems are designed…

人机交互 · 计算机科学 2024-09-20 Zhipeng Li , Christoph Gebhardt , Yves Inglin , Nicolas Steck , Paul Streli , Christian Holz

Conventional methods for visual assessment of civil infrastructures have certain limitations, such as subjectivity of the collected data, long inspection time, and high cost of labor. Although some new technologies i.e. robotic techniques…

计算机视觉与模式识别 · 计算机科学 2019-02-21 Enes Karaaslan , Ulas Bagci , F. Necati Catbas

In recent years, multimodal large language models (MLLMs) have shown remarkable capabilities in tasks like visual question answering and common sense reasoning, while visual perception models have made significant strides in perception…

计算机视觉与模式识别 · 计算机科学 2024-06-25 Guanqun Wang , Xinyu Wei , Jiaming Liu , Ray Zhang , Yichi Zhang , Kevin Zhang , Maurice Chong , Shanghang Zhang

Prior studies report that partial driving automation can increase the cognitive demands on human drivers. This effect largely arises from human drivers' lack of transparent insight into the vehicle's intentions and decision logic, as well…

This paper introduces a software architecture for real-time object detection using machine learning (ML) in an augmented reality (AR) environment. Our approach uses the recent state-of-the-art YOLOv8 network that runs onboard on the…

计算机视觉与模式识别 · 计算机科学 2023-06-07 Mikołaj Łysakowski , Kamil Żywanowski , Adam Banaszczyk , Michał R. Nowicki , Piotr Skrzypczyński , Sławomir K. Tadeja

Mixed reality headsets, such as the Microsoft HoloLens 2, are powerful sensing devices with integrated compute capabilities, which makes it an ideal platform for computer vision research. In this technical report, we present HoloLens 2…

Attention-deficit/hyperactivity disorder (ADHD) is a neurodevelopmental disorder that is highly prevalent and requires clinical specialists to diagnose. It is known that an individual's viewing behavior, reflected in their eye movements, is…

During collaboration in XR (eXtended Reality), users typically share and interact with virtual objects in a common, shared virtual environment. Specifically, collaboration among users in Mixed Reality (MR) requires knowing their position,…

人机交互 · 计算机科学 2023-10-10 Hung-Jui Guo , Omeed Eshaghi Ashtiani , Balakrishnan Prabhakaran

Mixed Reality (MR) devices are being increasingly adopted across a wide range of real-world applications, ranging from education and healthcare to remote work and entertainment. However, the unique immersive features of MR devices, such as…

人机交互 · 计算机科学 2025-01-29 Maha Sajid , Syed Ibrahim Mustafa Shah Bukhari , Bo Ji , Brendan David-John

Mixed Reality enables hybrid workspaces where physical and virtual monitors are adaptively created and moved to suit the current environment and needs. However, in shared settings, individual users' workspaces are rarely aligned and can…

人机交互 · 计算机科学 2025-08-14 Ludwig Sidenmark , Tianyu Zhang , Leen Al Lababidi , Jiannan Li , Tovi Grossman

Independent life of the individuals suffering from Alzheimer's disease (AD) is compromised due to their memory loss. As a result, they depend on others to help them lead their daily life. In this situation, either the family members or the…

人机交互 · 计算机科学 2024-03-12 Fatemeh Ghorbani

We tackle the problem of understanding visual ads where given an ad image, our goal is to rank appropriate human generated statements describing the purpose of the ad. This problem is generally addressed by jointly embedding images and…

计算机视觉与模式识别 · 计算机科学 2018-07-05 Karuna Ahuja , Karan Sikka , Anirban Roy , Ajay Divakaran

Robotic ultrasound systems can enhance medical diagnostics, but patient acceptance is a challenge. We propose a system combining an AI-powered conversational virtual agent with three mixed reality visualizations to improve trust and…

人机交互 · 计算机科学 2025-07-18 Tianyu Song , Felix Pabst , Ulrich Eck , Nassir Navab

Depression has been the leading cause of mental-health illness worldwide. Major depressive disorder (MDD), is a common mental health disorder that affects both psychologically as well as physically which could lead to loss of lives. Due to…

计算机视觉与模式识别 · 计算机科学 2019-09-05 Anupama Ray , Siddharth Kumar , Rutvik Reddy , Prerana Mukherjee , Ritu Garg

Independently exploring unknown spaces or finding objects in an indoor environment is a daily but challenging task for visually impaired people. However, common 2D assistive systems lack depth relationships between various objects,…

计算机视觉与模式识别 · 计算机科学 2021-07-08 Huayao Liu , Ruiping Liu , Kailun Yang , Jiaming Zhang , Kunyu Peng , Rainer Stiefelhagen

In this article, we introduce a novel problem of audio-visual autism behavior recognition, which includes social behavior recognition, an essential aspect previously omitted in AI-assisted autism screening research. We define the task at…

Multimodal Stance Detection (MSD) is crucial for understanding public discourse, yet effectively fusing text and image, especially with conflicting signals, remains challenging. Existing methods often face difficulties with contextual…

人工智能 · 计算机科学 2026-05-01 Weihai Lu , Zhejun Zhao , Yanshu Li , Huan He