English
Related papers

Related papers: XiCAD: Camera Activation Detection in the Da Vinci…

200 papers

Comprehensive perception of human beings is the prerequisite to ensure the safety of human-robot interaction. Currently, prevailing visual sensing approach typically involves a single static camera, resulting in a restricted and occluded…

Robotics · Computer Science 2024-03-20 Yuanjiong Ying , Xian Huang , Wei Dong

Purpose: Foundation models, trained on multitudes of public datasets, often require additional fine-tuning or re-prompting mechanisms to be applied to visually distinct target domains such as surgical videos. Further, without domain…

Image and Video Processing · Electrical Eng. & Systems 2025-07-02 Ssharvien Kumar Sivakumar , Yannik Frisch , Amin Ranem , Anirban Mukhopadhyay

Purpose: The objective of this investigation is to provide a comprehensive analysis of state-of-the-art methods for video-based assessment of surgical skill in the operating room. Methods: Using a data set of 99 videos of capsulorhexis, a…

Computer Vision and Pattern Recognition · Computer Science 2022-05-16 Sanchit Hira , Digvijay Singh , Tae Soo Kim , Shobhit Gupta , Gregory Hager , Shameema Sikder , S. Swaroop Vedula

A key element of computer-assisted surgery systems is phase recognition of surgical videos. Existing phase recognition algorithms require frame-wise annotation of a large number of videos, which is time and money consuming. In this work we…

Computer Vision and Pattern Recognition · Computer Science 2023-10-27 Roy Hirsch , Regev Cohen , Mathilde Caron , Tomer Golany , Daniel Freedman , Ehud Rivlin

Purpose: Drop-in gamma probes are widely used in robotic-assisted minimally invasive surgery (RAMIS) for lymph node detection. However, these devices only provide audio feedback on signal intensity, lacking the visual feedback necessary for…

Image and Video Processing · Electrical Eng. & Systems 2024-10-31 Songyu Xu , Yicheng Hu , Jionglong Su , Daniel Elson , Baoru Huang

Video action detection requires dense spatio-temporal annotations, which are both challenging and expensive to obtain. However, real-world videos often vary in difficulty and may not require the same level of annotation. This paper analyzes…

Computer Vision and Pattern Recognition · Computer Science 2025-08-20 Aayush Rana , Akash Kumar , Vibhav Vineet , Yogesh S Rawat

Under conventional "open-" surgery, the physician has to take care of the patient, interact with other clinicians and check several monitoring devices. Nowadays, the Computer Assisted Surgery proposes to integrate 3D cameras in the…

Medical Physics · Physics 2016-08-16 Jose Vázquez-Buenosaires , Yohan Payan , Jacques Demongeot

Objective. This paper presents an overview of generalizable and explainable artificial intelligence (XAI) in deep learning (DL) for medical imaging, aimed at addressing the urgent need for transparency and explainability in clinical…

Computer Vision and Pattern Recognition · Computer Science 2025-03-12 Ahmad Chaddad , Yan Hu , Yihang Wu , Binbin Wen , Reem Kateb

In human-robot interaction (HRI), detecting a human's gaze helps robots interpret user attention and intent. However, most gaze detection approaches rely on specialized eye-tracking hardware, limiting deployment in everyday settings.…

Robotics · Computer Science 2026-03-18 Linlin Cheng , Koen Hindriks , Artem V. Belopolsky

Robotic automation in surgery requires precise tracking of surgical tools and mapping of deformable tissue. Previous works on surgical perception frameworks require significant effort in developing features for surgical tool and tissue…

Robotics · Computer Science 2021-03-26 Jingpei Lu , Ambareesh Jayakumari , Florian Richter , Yang Li , Michael C. Yip

In this report, we present the technical details of our approach to the EPIC-KITCHENS-100 Unsupervised Domain Adaptation (UDA) Challenge for Action Recognition. The EPIC-KITCHENS-100 dataset consists of daily kitchen activities focusing on…

Computer Vision and Pattern Recognition · Computer Science 2022-06-07 Yi Cheng , Fen Fang , Ying Sun

Interactive image segmentation aims at obtaining a segmentation mask for an image using simple user annotations. During each round of interaction, the segmentation result from the previous round serves as feedback to guide the user's…

Computer Vision and Pattern Recognition · Computer Science 2023-03-22 Qiaoqiao Wei , Hui Zhang , Jun-Hai Yong

In this project, and through an understanding of neuronal system communication, A novel model serves as an assistive technology for locked-in people suffering from Motor neuronal disease (MND) is proposed. Work was done upon the potential…

Medical Physics · Physics 2018-09-05 Mahmoud Haroun , Mohamed Salah

This study examined whether a single ceiling-mounted camera could be used to capture fine-grained learning behaviours in co-located practical learning. In undergraduate nursing simulations, teachers first identified seven observable…

Human-Computer Interaction · Computer Science 2026-03-17 Xinyu Li , Linxuan Zhao , Roberto Martinez-Maldonado , Dragan Gasevic , Lixiang Yan

Micro-expression analysis has applications in domains such as Human-Robot Interaction and Driver Monitoring Systems. Accurately capturing subtle and fast facial movements remains difficult when relying solely on RGB cameras, due to…

Computer Vision and Pattern Recognition · Computer Science 2025-08-22 Nicolas Mastropasqua , Ignacio Bugueno-Cordova , Rodrigo Verschae , Daniel Acevedo , Pablo Negri , Maria E. Buemi

To the best of our knowledge, the existing deep-learning-based Video Super-Resolution (VSR) methods exclusively make use of videos produced by the Image Signal Processor (ISP) of the camera system as inputs. Such methods are 1) inherently…

Image and Video Processing · Electrical Eng. & Systems 2021-02-24 Xiaohong Liu , Kangdi Shi , Zhe Wang , Jun Chen

Introduction: Technical burdens and time-intensive review processes limit the practical utility of video capsule endoscopy (VCE). Artificial intelligence (AI) is poised to address these limitations, but the intersection of AI and VCE…

Though action recognition in videos has achieved great success recently, it remains a challenging task due to the massive computational cost. Designing lightweight networks is a possible solution, but it may degrade the recognition…

Computer Vision and Pattern Recognition · Computer Science 2020-02-11 Wenhao Wu , Dongliang He , Xiao Tan , Shifeng Chen , Yi Yang , Shilei Wen

Human-Object Interaction (HOI) detection aims to identify humans and objects within images and interpret their interactions. Existing HOI methods rely heavily on large datasets with manual annotations to learn interactions from visual cues.…

Computer Vision and Pattern Recognition · Computer Science 2025-07-24 Francesco Tonini , Lorenzo Vaquero , Alessandro Conti , Cigdem Beyan , Elisa Ricci

The goal of spatial-temporal action detection is to determine the time and place where each person's action occurs in a video and classify the corresponding action category. Most of the existing methods adopt fully-supervised learning,…

Computer Vision and Pattern Recognition · Computer Science 2023-09-21 Wei-Jhe Huang , Jheng-Hsien Yeh , Min-Hung Chen , Gueter Josmy Faure , Shang-Hong Lai
‹ Prev 1 8 9 10 Next ›