中文
相关论文

相关论文: Know your sensORs -- A Modality Study For Surgical…

200 篇论文

Objectives Computer vision (CV) is a field of artificial intelligence that enables machines to interpret and understand images and videos. CV has the potential to be of assistance in the operating room (OR) to track surgical instruments. We…

Recognition of surgical gesture is crucial for surgical skill assessment and efficient surgery training. Prior works on this task are based on either variant graphical models such as HMMs and CRFs, or deep learning models such as Recurrent…

计算机视觉与模式识别 · 计算机科学 2018-06-22 Daochang Liu , Tingting Jiang

Spatiotemporal action recognition deals with locating and classifying actions in videos. Motivated by the latest state-of-the-art real-time object detector You Only Watch Once (YOWO), we aim to modify its structure to increase action…

计算机视觉与模式识别 · 计算机科学 2020-12-16 Shentong Mo , Xiaoqing Tan , Jingfei Xia , Pinxu Ren

The advent of telemedicine represents a transformative development in leveraging technology to extend the reach of specialized medical expertise to remote surgeries, a field where the immediacy of expert guidance is paramount. However, the…

图像与视频处理 · 电气工程与系统科学 2024-07-30 Yixuan Wu , Kaiyuan Hu , Qian Shao , Jintai Chen , Danny Z. Chen , Jian Wu

Kinematic trajectories recorded from surgical robots contain information about surgical gestures and potentially encode cues about surgeon's skill levels. Automatic segmentation of these trajectories into meaningful action units could help…

机器人学 · 计算机科学 2019-07-26 Beatrice van Amsterdam , Hirenkumar Nakawala , Elena De Momi , Danail Stoyanov

Human actions often involve complex interactions across several inter-related objects in the scene. However, existing approaches to fine-grained video understanding or visual relationship detection often rely on single object representation…

计算机视觉与模式识别 · 计算机科学 2018-03-22 Chih-Yao Ma , Asim Kadav , Iain Melvin , Zsolt Kira , Ghassan AlRegib , Hans Peter Graf

Visual-Language Models (VLMs) have significantly advanced action video recognition. Supervised by the semantics of action labels, recent works adapt the visual branch of VLMs to learn video representations. Despite the effectiveness proved…

计算机视觉与模式识别 · 计算机科学 2023-10-11 Yifei Chen , Dapeng Chen , Ruijin Liu , Hao Li , Wei Peng

To assist surgeons in the operating theatre, surgical phase recognition is critical for developing computer-assisted surgical systems, which requires comprehensive understanding of surgical videos. Although existing studies made great…

计算机视觉与模式识别 · 计算机科学 2023-11-23 Zhen Chen , Yuhao Zhai , Jun Zhang , Jinqiao Wang

Tactile gesture recognition systems play a crucial role in Human-Robot Interaction (HRI) by enabling intuitive communication between humans and robots. The literature mainly addresses this problem by applying machine learning techniques to…

机器人学 · 计算机科学 2025-08-07 Shaohong Zhong , Alessandro Albini , Giammarco Caroleo , Giorgio Cannata , Perla Maiolino

Object segmentation plays an important role in the modern medical image analysis, which benefits clinical study, disease diagnosis, and surgery planning. Given the various modalities of medical images, the automated or semi-automated…

图像与视频处理 · 电气工程与系统科学 2020-06-01 Dong Yang , Holger Roth , Xiaosong Wang , Ziyue Xu , Andriy Myronenko , Daguang Xu

Robot-assisted minimally invasive surgeries offer many advantages but require complex motor tasks that take surgeons years to master. There is currently a lack of knowledge on how surgeons acquire these robotic surgical skills. Toward…

机器人学 · 计算机科学 2026-01-05 Hanna Kossowsky Lev , Yarden Sharon , Alex Geftler , Ilana Nisky

Human Action Recognition (HAR) is a challenging domain in computer vision, involving recognizing complex patterns by analyzing the spatiotemporal dynamics of individuals' movements in videos. These patterns arise in sequential data, such as…

计算机视觉与模式识别 · 计算机科学 2025-01-23 Ali K. AlShami , Ryan Rabinowitz , Khang Lam , Yousra Shleibik , Melkamu Mersha , Terrance Boult , Jugal Kalita

We propose a new benchmark for evaluating stereoscopic visual-inertial computer vision algorithms (SLAM/ SfM/ 3D Reconstruction/ Visual-Inertial Odometry) for minimally invasive surgical (MIS) interventions in the abdomen. Our MITI Dataset…

图像与视频处理 · 电气工程与系统科学 2026-02-12 Regine Hartwig , Daniel Ostler , Jean-Claude Rosenthal , Hubertus Feußner , Dirk Wilhelm , Dirk Wollherr

Zero-Shot Action Recognition has attracted attention in the last years and many approaches have been proposed for recognition of objects, events and actions in images and videos. There is a demand for methods that can classify instances…

计算机视觉与模式识别 · 计算机科学 2021-12-17 Valter Estevam , Helio Pedrini , David Menotti

In this paper, we present two video processing techniques for contact-less estimation of the Respiratory Rate (RR) of framed subjects. Due to the modest extent of movements related to respiration in both infants and adults, specific…

图像与视频处理 · 电气工程与系统科学 2022-11-30 Veronica Mattioli , Davide Alinovi , Gianluigi Ferrari , Francesco Pisani , Riccardo Raheli

Scene recognition is one of the basic problems in computer vision research with extensive applications in robotics. When available, depth images provide helpful geometric cues that complement the RGB texture information and help to identify…

计算机视觉与模式识别 · 计算机科学 2021-09-08 Andrea Ferreri , Silvia Bucci , Tatiana Tommasi

Recording the open surgery process is essential for educational and medical evaluation purposes; however, traditional single-camera methods often face challenges such as occlusions caused by the surgeon's head and body, as well as…

计算机视觉与模式识别 · 计算机科学 2025-04-10 Xinyu Liu , Xiaoguang Lin , Xiang Liu , Yong Yang , Hongqian Wang , Qilong Sun

Existing methods on video-based action recognition are generally view-dependent, i.e., performing recognition from the same views seen in the training data. We present a novel multiview spatio-temporal AND-OR graph (MST-AOG) representation…

计算机视觉与模式识别 · 计算机科学 2014-05-14 Jiang wang , Xiaohan Nie , Yin Xia , Ying Wu , Song-Chun Zhu

Single modality action recognition on RGB or depth sequences has been extensively explored recently. It is generally accepted that each of these two modalities has different strengths and limitations for the task of action recognition.…

计算机视觉与模式识别 · 计算机科学 2016-12-28 Amir Shahroudy , Tian-Tsong Ng , Yihong Gong , Gang Wang

A modern operating room (OR) provides a plethora of advanced medical devices. In order to better facilitate the information offered by them, they need to automatically react to the intra-operative context. To this end, the progress of the…

机器学习 · 计算机科学 2017-06-05 Ralf Stauder , Ergün Kayis , Nassir Navab
‹ 上一页 1 8 9 10 下一页 ›