中文
相关论文

相关论文: Adding Safety Rules to Surgeon-Authored VR Trainin…

200 篇论文

Grasping is a fundamental task in robot-assisted surgery (RAS), and automating it can reduce surgeon workload while enhancing efficiency, safety, and consistency beyond teleoperated systems. Most prior approaches rely on explicit object…

机器人学 · 计算机科学 2025-08-18 Hongbin Lin , Bin Li , Kwok Wai Samuel Au

We propose Strongly Supervised pre-training with ScreenShots (S4) - a novel pre-training paradigm for Vision-Language Models using data from large-scale web screenshot rendering. Using web screenshots unlocks a treasure trove of visual and…

计算机视觉与模式识别 · 计算机科学 2025-03-14 Yuan Gao , Kunyu Shi , Pengkai Zhu , Edouard Belval , Oren Nuriel , Srikar Appalaraju , Shabnam Ghadar , Vijay Mahadevan , Zhuowen Tu , Stefano Soatto

In this contribution, we design, implement and evaluate the pedagogical benefits of a novel interactive note taking interface (iVRNote) in VR for the purpose of learning and reflection lectures. In future VR learning environments, students…

人机交互 · 计算机科学 2019-10-04 Yi-Ting Chen , Chi-Hsuan Hsu , Chih-Han Chung , Yu-Shuen Wang , Sabarish V. Babu

Surface visualizations are essential in analyzing three-dimensional spatiotemporal phenomena. Given its ability to provide enhanced spatial perception and scene maneuverability, virtual reality (VR) is an essential medium for surface…

人机交互 · 计算机科学 2024-12-11 Hamza Afzaal , Usman Alim

Accurate segmentation of surgical instrument tip is an important task for enabling downstream applications in robotic surgery, such as surgical skill assessment, tool-tissue interaction and deformation modeling, as well as surgical…

计算机视觉与模式识别 · 计算机科学 2023-09-06 Jiaqi Liu , Yonghao Long , Kai Chen , Cheuk Hei Leung , Zerui Wang , Qi Dou

Dexterous manipulation of objects in virtual environments with our bare hands, by using only a depth sensor and a state-of-the-art 3D hand pose estimator (HPE), is challenging. While virtual environments are ruled by physics, e.g. object…

计算机视觉与模式识别 · 计算机科学 2020-08-10 Guillermo Garcia-Hernando , Edward Johns , Tae-Kyun Kim

We propose a novel multi-modal and multi-task architecture for simultaneous low level gesture and surgical task classification in Robot Assisted Surgery (RAS) videos.Our end-to-end architecture is based on the principles of a long…

计算机视觉与模式识别 · 计算机科学 2018-05-03 Duygu Sarikaya , Khurshid A. Guru , Jason J. Corso

The surgical usage of Mixed Reality (MR) has received growing attention in areas such as surgical navigation systems, skill assessment, and robot-assisted surgeries. For such applications, pose estimation for hand and surgical instruments…

计算机视觉与模式识别 · 计算机科学 2023-10-05 Rui Wang , Sophokles Ktistakis , Siwei Zhang , Mirko Meboldt , Quentin Lohmeyer

As part of their training all medical students and residents have to pass basic surgical tasks such as knot tying, needle-passing, and suturing. Their assessment is typically performed in the operating room by surgical faculty where…

计算机视觉与模式识别 · 计算机科学 2023-12-27 Yunzhe Xue , Olanrewaju Eletta , Justin W. Ady , Nell M. Patel , Advaith Bongu , Usman Roshan

Building assistive interfaces for controlling robots through arbitrary, high-dimensional, noisy inputs (e.g., webcam images of eye gaze) can be challenging, especially when it involves inferring the user's desired action in the absence of a…

机器人学 · 计算机科学 2022-02-08 Sean Chen , Jensen Gao , Siddharth Reddy , Glen Berseth , Anca D. Dragan , Sergey Levine

Personalized therapy, in which a therapeutic practice is adapted to an individual patient, can lead to improved health outcomes. Typically, this is accomplished by relying on a therapist's training and intuition along with feedback from a…

机器学习 · 计算机科学 2025-04-22 Athar Mahmoudi-Nejad , Matthew Guzdial , Pierre Boulanger

Assessing learner competency in clinical simulation requires expert observation that is time-intensive, difficult to scale, and subject to inter-rater variability. Vision-language models have emerged as a promising tool for understanding…

计算机视觉与模式识别 · 计算机科学 2026-05-21 Hanchen David Wang , Yilin Liu , Madison J. Lee , Surya Chand Rayala , Gautam Biswas , Daniel T. Levin , Meiyi Ma

Purpose: This research aims to facilitate the use of state-of-the-art computer vision algorithms for the automated training of surgeons and the analysis of surgical footage. By estimating 2D hand poses, we model the movement of the…

计算机视觉与模式识别 · 计算机科学 2023-04-03 Eddie Bkheet , Anne-Lise D'Angelo , Adam Goldbraikh , Shlomi Laufer

We investigate how vibrotactile wrist feedback can enhance spatial guidance for handheld tool movement in optical see-through augmented reality (AR). While AR overlays are widely used to support surgical tasks, visual occlusion, lighting…

人机交互 · 计算机科学 2026-01-21 Yue Yang , Christoph Leuze , Brian Hargreaves , Bruce Daniel , Fred M Baik

Computer-Assisted Intervention (CAI) has the potential to revolutionize modern surgery, with surgical scene understanding serving as a critical component in supporting decision-making, improving procedural efficacy, and ensuring…

Robot-assisted minimally invasive surgery is improving surgeon performance and patient outcomes. This innovation is also turning what has been a subjective practice into motion sequences that can be precisely measured. A growing number of…

计算机视觉与模式识别 · 计算机科学 2020-06-12 Neil Getty , Zixuan Zhao , Stephan Gruessner , Liaohai Chen , Fangfang Xia

The quality of Virtual Reality (VR) apps is vital, particularly the rendering quality of the VR Graphical User Interface (GUI). Different from traditional 2D apps, VR apps create a 3D digital scene for users, by rendering two distinct 2D…

软件工程 · 计算机科学 2024-10-30 Shuqing Li , Cuiyun Gao , Jianping Zhang , Yujia Zhang , Yepang Liu , Jiazhen Gu , Yun Peng , Michael R. Lyu

High-resolution optical tactile sensors are increasingly used in robotic learning environments due to their ability to capture large amounts of data directly relating to agent-environment interaction. However, there is a high barrier of…

机器人学 · 计算机科学 2022-07-28 Yijiong Lin , John Lloyd , Alex Church , Nathan F. Lepora

Text-to-speech (TTS) models have achieved remarkable naturalness in recent years, yet like most deep neural models, they have more parameters than necessary. Sparse TTS models can improve on dense models via pruning and extra retraining, or…

音频与语音处理 · 电气工程与系统科学 2024-06-04 Perry Lam , Huayun Zhang , Nancy F. Chen , Berrak Sisman , Dorien Herremans

VR Facial Animation is necessary in applications requiring clear view of the face, even though a VR headset is worn. In our case, we aim to animate the face of an operator who is controlling our robotic avatar system. We propose a real-time…

计算机视觉与模式识别 · 计算机科学 2023-04-25 Andre Rochow , Max Schwarz , Michael Schreiber , Sven Behnke