中文
相关论文

相关论文: Single- and Multi-Task Architectures for Tool Pres…

200 篇论文

Prior work demonstrated the ability of machine learning to automatically recognize surgical workflow steps from videos. However, these studies focused on only a single type of procedure. In this work, we analyze, for the first time,…

计算机视觉与模式识别 · 计算机科学 2021-04-22 Daniel Neimark , Omri Bar , Maya Zohar , Gregory D. Hager , Dotan Asselmann

One of the key challenges of detecting AI-generated images is spotting images that have been created by previously unseen generative models. We argue that the limited diversity of the training data is a major obstacle to addressing this…

计算机视觉与模式识别 · 计算机科学 2025-06-11 Jeongsoo Park , Andrew Owens

Analyzing laparoscopic surgery videos presents a complex and multifaceted challenge, with applications including surgical training, intra-operative surgical complication prediction, and post-operative surgical assessment. Identifying…

计算机视觉与模式识别 · 计算机科学 2023-12-04 Sahar Nasirihaghighi , Negin Ghamsarian , Heinrich Husslein , Klaus Schoeffmann

Cell detection and segmentation is fundamental for all downstream analysis of digital pathology images. However, obtaining the pixel-level ground truth for single cell segmentation is extremely labor intensive. To overcome this challenge,…

计算机视觉与模式识别 · 计算机科学 2019-10-29 Alireza Chamanzar , Yao Nie

This research paper presents an innovative multi-task learning framework that allows concurrent depth estimation and semantic segmentation using a single camera. The proposed approach is based on a shared encoder-decoder architecture, which…

计算机视觉与模式识别 · 计算机科学 2024-03-19 Pardis Taghavi , Reza Langari , Gaurav Pandey

The proliferation of deepfake imagery poses escalating challenges for practitioners tasked with verifying digital media authenticity. While detection algorithm research is abundant, empirical evaluations of publicly accessible tools that…

密码学与安全 · 计算机科学 2026-03-06 Michael Rettinger , Ben Beaumont , Nhien-An Le-Khac , Hong-Hanh Nguyen-Le

This work explores various ways of exploring multi-task learning (MTL) techniques aimed at classifying videos as original or manipulated in cross-manipulation scenario to attend generalizability in deep fake scenario. The dataset used in…

计算机视觉与模式识别 · 计算机科学 2023-08-28 Pranav Balaji , Abhijit Das , Srijan Das , Antitza Dantcheva

Surgical AI often involves multiple tasks within a single procedure, like phase recognition or assessing the Critical View of Safety in laparoscopic cholecystectomy. Traditional models, built for one task at a time, lack flexibility,…

计算机视觉与模式识别 · 计算机科学 2025-07-11 Soham Walimbe , Britty Baby , Vinkle Srivastav , Nicolas Padoy

Purpose: As visual inspection is an inherent process during radiological screening, the associated eye gaze data can provide valuable insights into relevant clinical decisions. As deep learning has become the state-of-the-art for…

图像与视频处理 · 电气工程与系统科学 2025-02-18 Zirui Qiu , Hassan Rivaz , Yiming Xiao

The amount of surgical data, recorded during video-monitored surgeries, has extremely increased. This paper aims at improving existing solutions for the automated analysis of cataract surgeries in real time. Through the analysis of a video…

计算机视觉与模式识别 · 计算机科学 2016-09-20 Hassan Al Hajj , Gwenolé Quellec , Mathieu Lamard , Guy Cazuguel , Béatrice Cochener

Recognizing the phases of a laparoscopic surgery (LS) operation form its video constitutes a fundamental step for efficient content representation, indexing and retrieval in surgical video databases. In the literature, most techniques focus…

计算机视觉与模式识别 · 计算机科学 2021-07-27 Constantinos Loukas

Instrument-tissue interaction detection task, which helps understand surgical activities, is vital for constructing computer-assisted surgery systems but with many challenges. Firstly, most models represent instrument-tissue interaction in…

计算机视觉与模式识别 · 计算机科学 2024-04-02 Wenjun Lin , Yan Hu , Huazhu Fu , Mingming Yang , Chin-Boon Chng , Ryo Kawasaki , Cheekong Chui , Jiang Liu

For several cancer patients, operative resection with curative intent can end up in early recurrence of the cancer. Current limitations in peri-operative cancer staging and especially intra-operative misidentification of visible metastases…

图像与视频处理 · 电气工程与系统科学 2023-06-21 Thomas Schnelldorfer , Janil Castro , Atoussa Goldar-Najafi , Liping Liu

We propose a new multilabel classifier, called LapTool-Net to detect the presence of surgical tools in each frame of a laparoscopic video. The novelty of LapTool-Net is the exploitation of the correlation among the usage of different tools…

计算机视觉与模式识别 · 计算机科学 2019-05-23 Babak Namazi , Ganesh Sankaranarayanan , Venkat Devarajan

Sustaining high fidelity and high throughput of perception tasks over vision sensor streams on edge devices remains a formidable challenge, especially given the continuing increase in image sizes (e.g., generated by 4K cameras) and…

多媒体 · 计算机科学 2023-05-08 Ila Gokarn , Hemanth Sabella , Yigong Hu , Tarek Abdelzaher , Archan Misra

Recognizing human actions in videos requires spatial and temporal understanding. Most existing action recognition models lack a balanced spatio-temporal understanding of videos. In this work, we propose a novel two-stream architecture,…

计算机视觉与模式识别 · 计算机科学 2024-09-04 Dongho Lee , Jongseo Lee , Jinwoo Choi

Machine learning, particularly convolutional neural networks (CNNs), has shown promise in medical image analysis, especially for thoracic disease detection using chest X-ray images. In this study, we evaluate various CNN architectures,…

计算机视觉与模式识别 · 计算机科学 2025-02-18 Tejas Mirthipati

The recognition of information in floor plan data requires the use of detection and segmentation models. However, relying on several single-task models can result in ineffective utilization of relevant information when there are multiple…

计算机视觉与模式识别 · 计算机科学 2023-09-04 Lingxiao Huang , Jung-Hsuan Wu , Chiching Wei , Wilson Li