中文
相关论文

相关论文: Multi-view Video-Pose Pretraining for Operating Ro…

200 篇论文

We present a novel method for intraoperative patient-to-image registration by learning Expected Appearances. Our method uses preoperative imaging to synthesize patient-specific expected views through a surgical microscope for a predicted…

计算机视觉与模式识别 · 计算机科学 2023-10-04 Nazim Haouchine , Reuben Dorent , Parikshit Juvekar , Erickson Torio , William M. Wells , Tina Kapur , Alexandra J. Golby , Sarah Frisken

This work addresses the problem of Social Activity Recognition (SAR), a critical component in real-world tasks like surveillance and assistive robotics. Unlike traditional event understanding approaches, SAR necessitates modeling individual…

计算机视觉与模式识别 · 计算机科学 2024-06-21 Shubham Trehan , Sathyanarayanan N. Aakur

Detection of surgical instruments plays a key role in ensuring patient safety in minimally invasive surgery. In this paper, we present a novel method for 2D vision-based recognition and pose estimation of surgical instruments that…

计算机视觉与模式识别 · 计算机科学 2017-10-19 Thomas Kurmann , Pablo Marquez Neila , Xiaofei Du , Pascal Fua , Danail Stoyanov , Sebastian Wolf , Raphael Sznitman

In the realm of automated robotic surgery and computer-assisted interventions, understanding robotic surgical activities stands paramount. Existing algorithms dedicated to surgical activity recognition predominantly cater to pre-defined…

计算机视觉与模式识别 · 计算机科学 2024-02-13 Long Bai , Guankun Wang , Jie Wang , Xiaoxiao Yang , Huxin Gao , Xin Liang , An Wang , Mobarakol Islam , Hongliang Ren

Endoscopic surgery is the gold standard for robotic-assisted minimally invasive surgery, offering significant advantages in early disease detection and precise interventions. However, the complexity of surgical scenes, characterized by high…

计算机视觉与模式识别 · 计算机科学 2025-06-10 Guankun Wang , Rui Tang , Mengya Xu , Long Bai , Huxin Gao , Hongliang Ren

We propose the use of self-supervised learning for human activity recognition with smartphone accelerometer data. Our proposed solution consists of two steps. First, the representations of unlabeled input signals are learned by training a…

信号处理 · 电气工程与系统科学 2021-09-03 Setareh Rahimi Taghanaki , Michael Rainbow , Ali Etemad

Surgical robotics holds much promise for improving patient safety and clinician experience in the Operating Room (OR). However, it also comes with new challenges, requiring strong team coordination and effective OR management. Automatic…

计算机视觉与模式识别 · 计算机科学 2023-12-20 Idris Hamoud , Muhammad Abdullah Jamal , Vinkle Srivastav , Didier Mutter , Nicolas Padoy , Omid Mohareri

Robust and accurate 2D/3D registration, which aligns preoperative models with intraoperative images of the same anatomy, is crucial for successful interventional navigation. To mitigate the challenge of a limited field of view in…

计算机视觉与模式识别 · 计算机科学 2025-06-30 Yuxin Cui , Rui Song , Yibin Li , Max Q. -H. Meng , Zhe Min

Human action recognition (HAR) in videos is one of the core tasks of video understanding. Based on video sequences, the goal is to recognize actions performed by humans. While HAR has received much attention in the visible spectrum, action…

计算机视觉与模式识别 · 计算机科学 2022-04-20 Soufiane Lamghari , Guillaume-Alexandre Bilodeau , Nicolas Saunier

Real-time intelligent detection and prediction of subjects' behavior particularly their movements or actions is critical in the ward. This approach offers the advantage of reducing in-hospital care costs and improving the efficiency of…

计算机视觉与模式识别 · 计算机科学 2023-10-06 Zherui Li , Raye Chen-Hua Yeow

Video Action Recognition (VAR) is a challenging task due to its inherent complexities. Though different approaches have been explored in the literature, designing a unified framework to recognize a large number of human actions is still a…

计算机视觉与模式识别 · 计算机科学 2023-08-09 Soumyabrata Chaudhuri , Saumik Bhattacharya

Human activity recognition (HAR) with deep learning models relies on large amounts of labeled data, often challenging to obtain due to associated cost, time, and labor. Self-supervised learning (SSL) has emerged as an effective approach to…

计算机视觉与模式识别 · 计算机科学 2025-03-25 Dominique Nshimyimana , Vitor Fortes Rey , Sungho Suh , Bo Zhou , Paul Lukowicz

Skeleton-based human action recognition is a longstanding challenge due to its complex dynamics. Some fine-grain details of the dynamics play a vital role in classification. The existing work largely focuses on designing incremental neural…

计算机视觉与模式识别 · 计算机科学 2023-09-06 Ruijie Hou , Yanran Li , Ningyu Zhang , Yulin Zhou , Xiaosong Yang , Zhao Wang

Most recent view-invariant action recognition and performance assessment approaches rely on a large amount of annotated 3D skeleton data to extract view-invariant features. However, acquiring 3D skeleton data can be cumbersome, if not…

计算机视觉与模式识别 · 计算机科学 2024-07-09 Faegheh Sardari , Björn Ommer , Majid Mirmehdi

Person detection and pose estimation is a key requirement to develop intelligent context-aware assistance systems. To foster the development of human pose estimation methods and their applications in the Operating Room (OR), we release the…

计算机视觉与模式识别 · 计算机科学 2021-08-23 Vinkle Srivastav , Thibaut Issenhuth , Abdolrahim Kadkhodamohammadi , Michel de Mathelin , Afshin Gangi , Nicolas Padoy

Purpose: Microsurgical Aneurysm Clipping Surgery (MACS) carries a high risk for intraoperative aneurysm rupture. Automated recognition of instances when the aneurysm is exposed in the surgical video would be a valuable reference point for…

计算机视觉与模式识别 · 计算机科学 2023-03-20 Jinfan Zhou , William Muirhead , Simon C. Williams , Danail Stoyanov , Hani J. Marcus , Evangelos B. Mazomenos

Robustness to domain changes is a key capability for effective deployment of human action recognition systems in real-world scenarios, where action categories at inference can present important domain shifts or even unseen actions from…

计算机视觉与模式识别 · 计算机科学 2026-05-22 Yannick Porto , Renato Martins , Thomas Chalumeau , Cedric Demonceaux

EasyVis2 is a system designed to provide hands-free, real-time 3D visualization for laparoscopic surgery. It incorporates a surgical trocar equipped with an array of micro-cameras, which can be inserted into the body cavity to offer an…

计算机视觉与模式识别 · 计算机科学 2025-04-10 Yung-Hong Sun , Gefei Shen , Jiangang Chen , Jayer Fernandes , Amber L. Shada , Charles P. Heise , Hongrui Jiang , Yu Hen Hu

Kinematic rigs provide a structured interface for articulating 3D meshes but lack any associated pose space, i.e., an explicit representation of the plausible manifold of joint configurations for a given mesh. Without such a pose space,…

计算机视觉与模式识别 · 计算机科学 2026-05-22 Honglin Chen , Karran Pandey , Rundi Wu , Matheus Gadelha , Yannick Hold-Geoffroy , Ayush Tewari , Niloy J. Mitra , Changxi Zheng , Paul Guerrero

Mistake analysis in procedural activities is a critical area of research with applications spanning industrial automation, physical rehabilitation, education and human-robot collaboration. This paper reviews vision-based methods for…

计算机视觉与模式识别 · 计算机科学 2025-12-04 Konstantinos Bacharidis , Antonis A. Argyros