中文
相关论文

相关论文: A Virtual Reality Game to Improve Physical and Cog…

200 篇论文

While visual data augmentation remains a cornerstone for training robust vision models, it has received limited attention in visual language models (VLMs), which predominantly rely on large-scale real data acquisition or synthetic…

计算机视觉与模式识别 · 计算机科学 2025-12-03 Zhengzhuo Xu , Chong Sun , SiNan Du , Chen Li , Jing Lyu , Chun Yuan

The current study examines how adequate coordination among different cognitive processes including visual recognition, attention switching, action preparation and generation can be developed via learning of robots by introducing a novel…

人工智能 · 计算机科学 2016-11-15 Jungsik Hwang , Minju Jung , Naveen Madapana , Jinhyung Kim , Minkyu Choi , Jun Tani

Vision-language reinforcement learning (RL) has primarily focused on narrow domains (e.g. geometry or chart reasoning). This leaves broader training scenarios and resources underexplored, limiting the exploration and learning of Vision…

Pseudo-haptic techniques are used to modify haptic perception by appropriately changing visual feedback to body movements. Based on the knowledge that tendon vibration can affect our somatosensory perception, this paper proposes a method…

人机交互 · 计算机科学 2023-08-31 Yutaro Hirao , Tomohiro Amemiya , Takuji Narumi , Ferran Argelaguet , Anatole Lécuyer

Visual instruction tuning is a key training stage of large multimodal models. However, when learning multiple visual tasks simultaneously, this approach often results in suboptimal and imbalanced overall performance due to latent knowledge…

人工智能 · 计算机科学 2026-01-22 Yanqi Dai , Yong Wang , Zebin You , Dong Jing , Xiangxiang Chu , Zhiwu Lu

Color Vision Deficiency (CVD) affects nearly 8 percent of men and 0.5 percent of women worldwide. Existing color-correction methods often rely on prior clinical diagnosis and static filtering, making them less effective for users with mild…

人机交互 · 计算机科学 2025-09-11 Jingwen Qin , Semen Checherin , Yue Li , Berend-Jan van der Zwaag , Ozlem Durmaz-Incel

Recent advances in Vision-Language Models (VLMs) have substantially enhanced their ability across multimodal video understanding benchmarks spanning temporal, action, object, and spatial understanding. However, we identify a critical yet…

计算机视觉与模式识别 · 计算机科学 2026-04-21 Cui Yakun , Xingqun Qi , TianTian Geng , Yuyao Zhang , Sirui Han , Yike Guo

Existing methods of haptic feedback for virtual fluids are challenging to scale, lack durability for long-term rough use, and fail to fully capture the expressive haptic qualities of fluids. To overcome these limitations, we present…

人机交互 · 计算机科学 2025-02-03 Frank Wencheng Liu , Ryan Wirjadi , Yanjun Lyu , Shiling Dai , Byron Lahey , Assegid Kidane , Robert LiKamWa

We aim to ask and answer an essential question "how quickly do we react after observing a displayed visual target?" To this end, we present psychophysical studies that characterize the remarkable disconnect between human saccadic behaviors…

人机交互 · 计算机科学 2022-05-06 Budmonde Duinkharjav , Praneeth Chakravarthula , Rachel Brown , Anjul Patney , Qi Sun

Increasing individuals' awareness of their own body signals can lead to improved interoception, enabling the brain to estimate current body states more accurately and in a timely manner. However, certain body signals, such as eye movements,…

人机交互 · 计算机科学 2023-10-23 Songlin Xu , Xinyu Zhang

Recent progress in artificial intelligence through reinforcement learning (RL) has shown great success on increasingly complex single-agent environments and two-player turn-based games. However, the real-world contains multiple agents, each…

The game Quantum Moves was designed to pit human players against computer algorithms, combining their solutions into hybrid optimization to control a scalable quantum computer. In this midstream report, we open our design process and…

计算机与社会 · 计算机科学 2015-06-30 Andreas Lieberoth , Mads Kock Pedersen , Andreea Catalina Marin , Tilo Planke , Jacob Friis Sherson

Traditional cybersecurity tabletop exercises (TTXs) provide valuable training but are often scripted, resource-intensive, and difficult to scale. We introduce AgentBnB, a browser-based re-imagining of the Backdoors & Breaches game that…

计算与语言 · 计算机科学 2025-11-04 Arman Anwar , Zefang Liu

Optical see-through augmented reality (OST-AR) overlays digital targets and annotations on the physical world, offering promising guidance for hands-on tasks such as medical needle insertion or assembly. Recent work on OST-AR depth…

图形学 · 计算机科学 2025-10-03 Hu Guo , Lily Patel , Rohan Gupt

The hubness problem widely exists in high-dimensional embedding space and is a fundamental source of error for cross-modal matching tasks. In this work, we study the emergence of hubs in Visual Semantic Embeddings (VSE) with application to…

机器学习 · 计算机科学 2019-11-25 Fangyu Liu , Rongtian Ye , Xun Wang , Shuaipeng Li

We propose GAM-Agent, a game-theoretic multi-agent framework for enhancing vision-language reasoning. Unlike prior single-agent or monolithic models, GAM-Agent formulates the reasoning process as a non-zero-sum game between base…

人工智能 · 计算机科学 2025-05-30 Jusheng Zhang , Yijia Fan , Wenjun Lin , Ruiqi Chen , Haoyi Jiang , Wenhao Chai , Jian Wang , Keze Wang

Virtual reality (VR) is not a new technology but has been in development for decades, driven by advances in computer technology. Currently, VR technology is increasingly being used in applications to enable immersive, yet controlled…

人机交互 · 计算机科学 2022-11-24 Hong Gao

Many graphics rendering algorithms used in both real-time games and virtual reality applications can get performance boosts by temporally reusing previous computations. However, algorithms based on temporal reuse are typically measured…

图形学 · 计算机科学 2023-05-09 Erfan Momeni Yazdi , Markku Mäkitalo , Julius Ikkala , Pekka Jääskeläinen

Surgical simulators have been widely used in training and evaluation of physicians and surgeons. Virtual reality augmented with haptic technology has made it feasible to develop more realistic surgical simulators. In this context, we set…

医学物理 · 物理学 2022-05-19 Reza Karimzadeh , Javad Sheikh , Hamed Azarnoush , Hossein Arabi

In order to engage in complex social interaction, humans learn at a young age to infer what others see and cannot see from a different point-of-view, and learn to predict others' plans and behaviors. These abilities have been mostly lacking…

机器人学 · 计算机科学 2021-05-12 Boyuan Chen , Yuhang Hu , Robert Kwiatkowski , Shuran Song , Hod Lipson