中文
相关论文

相关论文: 2024 NASA SUITS Report: LLM-Driven Immersive Augme…

200 篇论文

Multimodal large language models (MLLMs) have shown remarkable capabilities in cross-modal understanding and reasoning, offering new opportunities for intelligent assistive systems, yet existing systems still struggle with risk-aware…

机器人学 · 计算机科学 2026-04-08 Renjun Gao

Unmanned Aerial Vehicles (UAVs) offer agile, secure and efficient solutions for communication relay networks. However, their modeling and control are challenging, and the mismatch between simulations and actual conditions limits real-world…

机器人学 · 计算机科学 2025-01-31 Yousef Emami , Kai Li , Luis Almeida , Sai Zou , Wei Ni

Virtual Reality (VR) interfaces are increasingly used as remote visualization media in telerobotics. Remote environments captured through RGB-D cameras and visualized using VR interfaces can enhance operators' situational awareness and…

人机交互 · 计算机科学 2022-10-12 Y. T. Tefera , D. Mazzanti , S. Anastasi , D. G. Caldwell , P. Fiorini , N. Deshpande

Extended Reality (XR) is a rapidly growing field offering unique immersive experiences, social networking, learning, and collaboration opportunities. The continuous advancements in XR technology and industry efforts are gradually moving…

人机交互 · 计算机科学 2023-08-23 Pascal Knierim , Thomas Kosch

Large language models (LLMs) have enabled the automatic generation of step-by-step augmented reality (AR) instructions for a wide range of physical tasks. However, existing LLM-based AR guidance often lacks rich visual augmentations to…

人机交互 · 计算机科学 2025-09-25 Ada Yi Zhao , Aditya Gunturu , Ellen Yi-Luen Do , Ryo Suzuki

Scaling up robot learning will likely require human data containing rich and long-horizon interactions in the wild. Existing approaches for collecting such data trade off portability, robustness to occlusion, and global consistency. We…

机器人学 · 计算机科学 2026-04-09 Wenjing Margaret Mao , Jefferson Ng , Luyang Hu , Daniel Gehrig , Antonio Loquercio

Establishing common ground between an intelligent robot and a human requires communication of the robot's intention, behavior, and knowledge to the human to build trust and assure safety in a shared environment. This paper introduces SENSAR…

机器人学 · 计算机科学 2020-11-10 Andre Cleaver , Faizan Muhammad , Amel Hassan , Elaine Short , Jivko Sinapov

Safe UAV emergency landing requires more than just identifying flat terrain; it demands understanding complex semantic risks (e.g., crowds, temporary structures) invisible to traditional geometric sensors. In this paper, we propose a novel…

计算机视觉与模式识别 · 计算机科学 2026-02-03 Chunliang Hua , Zeyuan Yang , Lei Zhang , Jiayang Sun , Fengwen Chen , Chunlan Zeng , Xiao Hu

Socially aware navigation is a fast-evolving research area in robotics that enables robots to move within human environments while adhering to the implicit human social norms. The advent of Deep Reinforcement Learning (DRL) has accelerated…

机器人学 · 计算机科学 2025-12-02 Ibrahim Khalil Kabir , Muhammad Faizan Mysorewala

In the evolving landscape of transportation systems, integrating Large Language Models (LLMs) offers a promising frontier for advancing intelligent decision-making across various applications. This paper introduces a novel 3-dimensional…

机器学习 · 计算机科学 2024-12-17 Dexter Le , Aybars Yunusoglu , Karn Tiwari , Murat Isik , I. Can Dikmen

Speaking aloud to a wearable AR assistant in public can be socially awkward, and re-articulating the same requests every day creates unnecessary effort. We present SpeechLess, a wearable AR assistant that introduces a speech-based intent…

人机交互 · 计算机科学 2026-04-14 Yoonsang Kim , Devshree Jadeja , Divyansh Pradhan , Yalong Yang , Arie Kaufman

The integration of Large Language Models (LLMs) into Virtual Reality (VR) games marks a paradigm shift in the design of immersive, adaptive, and intelligent digital experiences. This paper presents a comprehensive review of recent research…

人机交互 · 计算机科学 2025-11-24 Süeda Özkaya , Santiago Berrezueta-Guzman , Stefan Wagner

This paper introduces the Ambient Intelligence Rehabilitation Support (AIRS) framework, an advanced artificial intelligence-based solution tailored for home rehabilitation environments. AIRS integrates cutting-edge technologies, including…

人机交互 · 计算机科学 2025-07-14 Gábor Baranyi , Zsolt Csibi , Kristian Fenech , Áron Fóthi , Zsófia Gaál , Joul Skaf , András Lőrincz

Recent advancements in Large Language Models (LLMs) have catalyzed a paradigm shift from static prediction systems to agentic AI agents capable of reasoning, interacting with tools, and adapting to complex tasks. While LLM-based agentic…

计算机视觉与模式识别 · 计算机科学 2025-07-24 Nima Fathi , Amar Kumar , Tal Arbel

This work presents novel robot-mediated immersive experiences enabled by an encountered-type haptic display (ETHD) that introduces direct physical contact in virtual environments. We focus on social-physical interactions, a class of…

人机交互 · 计算机科学 2026-02-13 Eric Godden , Jacquie Groenewegen , Michael Wheeler , Matthew K. X. J. Pan

EXtended Reality (XR) is a rapidly developing paradigm for computer entertainment, and is also increasingly used for simulation, training, data analysis, and other non-entertainment purposes, often employing head-worn XR devices like the…

人机交互 · 计算机科学 2022-09-01 Thiago Porcino , Seyed Adel Ghaeinian , Juliano Franz , Joseph Malloch , Derek Reilly

To help improve the safety and accessibility of indoor spaces, researchers and health professionals have created assessment instruments that enable homeowners and trained experts to audit and improve homes. With advances in computer vision,…

人机交互 · 计算机科学 2022-10-07 Xia Su , Kaiming Cheng , Han Zhang , Jaewook Lee , Jon E. Froehlich

Drones have been widely used in many areas of our daily lives. It relieves people of the burden of holding a controller all the time and makes drone control easier to use for people with disabilities or occupied hands. However, the control…

计算机视觉与模式识别 · 计算机科学 2023-08-29 Xinyi Wang , Xuan Cui , Danxu Li , Fang Liu , Licheng Jiao

In robotics, Vision-Language-Action (VLA) models that integrate diverse multimodal signals from multi-view inputs have emerged as an effective approach. However, most prior work adopts static fusion that processes all visual inputs…

机器人学 · 计算机科学 2026-02-18 Young-Chae Son , Jung-Woo Lee , Yoon-Ji Choi , Dae-Kwan Ko , Soo-Chul Lim