中文
相关论文

相关论文: Single- and Multi-Task Architectures for Tool Pres…

200 篇论文

To assist surgeons in the operating theatre, surgical phase recognition is critical for developing computer-assisted surgical systems, which requires comprehensive understanding of surgical videos. Although existing studies made great…

计算机视觉与模式识别 · 计算机科学 2023-11-23 Zhen Chen , Yuhao Zhai , Jun Zhang , Jinqiao Wang

An abdominal ultrasound examination, which is the most common ultrasound examination, requires substantial manual efforts to acquire standard abdominal organ views, annotate the views in texts, and record clinically relevant organ…

计算机视觉与模式识别 · 计算机科学 2018-06-06 Zhoubing Xu , Yuankai Huo , JinHyeong Park , Bennett Landman , Andy Milkowski , Sasa Grbic , Shaohua Zhou

Seven million people suffer surgical complications each year, but with sufficient surgical training and review, 50\% of these complications could be prevented. To improve surgical performance, existing research uses various deep learning…

图像与视频处理 · 电气工程与系统科学 2022-04-19 Ella Selina Lan

In image-assisted minimally invasive surgeries (MIS), understanding surgical scenes is vital for real-time feedback to surgeons, skill evaluation, and improving outcomes through collaborative human-robot procedures. Within this context, the…

计算机视觉与模式识别 · 计算机科学 2024-12-13 Mithun Parab , Pranay Lendave , Jiyoung Kim , Thi Quynh Dan Nguyen , Palash Ingle

Human pose estimation and action recognition are related tasks since both problems are strongly dependent on the human body representation and analysis. Nonetheless, most recent methods in the literature handle the two problems separately.…

计算机视觉与模式识别 · 计算机科学 2020-03-05 Diogo C Luvizon , Hedi Tabia , David Picard

In this work, we propose an approach to the spatiotemporal localisation (detection) and classification of multiple concurrent actions within temporally untrimmed videos. Our framework is composed of three stages. In stage 1, appearance and…

计算机视觉与模式识别 · 计算机科学 2016-08-05 Suman Saha , Gurkirt Singh , Michael Sapienza , Philip H. S. Torr , Fabio Cuzzolin

Multi-task learning holds the promise of less data, parameters, and time than training of separate models. We propose a method to automatically search over multi-task architectures while taking resource constraints into consideration. We…

机器学习 · 计算机科学 2019-08-14 Alejandro Newell , Lu Jiang , Chong Wang , Li-Jia Li , Jia Deng

We propose a novel multi-modal and multi-task architecture for simultaneous low level gesture and surgical task classification in Robot Assisted Surgery (RAS) videos.Our end-to-end architecture is based on the principles of a long…

计算机视觉与模式识别 · 计算机科学 2018-05-03 Duygu Sarikaya , Khurshid A. Guru , Jason J. Corso

Multi-task learning improves generalization performance by sharing knowledge among related tasks. Existing models are for task combinations annotated on the same dataset, while there are cases where multiple datasets are available for each…

计算机视觉与模式识别 · 计算机科学 2018-05-16 Seiichiro Fukuda , Ryota Yoshihashi , Rei Kawakami , Shaodi You , Makoto Iida , Takeshi Naemura

This paper investigates the automatic monitoring of tool usage during a surgery, with potential applications in report generation, surgical training and real-time decision support. Two surgeries are considered: cataract surgery, the most…

计算机视觉与模式识别 · 计算机科学 2018-05-08 Hassan Al Hajj , Mathieu Lamard , Pierre-Henri Conze , Béatrice Cochener , Gwenolé Quellec

Introduction: Technical burdens and time-intensive review processes limit the practical utility of video capsule endoscopy (VCE). Artificial intelligence (AI) is poised to address these limitations, but the intersection of AI and VCE…

Purpose: Accurate assessment of surgical complexity is essential in Laparoscopic Cholecystectomy (LC), where severe inflammation is associated with longer operative times and increased risk of postoperative complications. The Parkland…

计算机视觉与模式识别 · 计算机科学 2025-11-07 Dimitrios Anastasiou , Santiago Barbarisi , Lucy Culshaw , Jayna Patel , Evangelos B. Mazomenos , Imanol Luengo , Danail Stoyanov

Sensor-based human activity segmentation and recognition are two important and challenging problems in many real-world applications and they have drawn increasing attention from the deep learning community in recent years. Most of the…

计算机视觉与模式识别 · 计算机科学 2023-03-21 Furong Duan , Tao Zhu , Jinqiang Wang , Liming Chen , Huansheng Ning , Yaping Wan

Lung cancer and covid-19 have one of the highest morbidity and mortality rates in the world. For physicians, the identification of lesions is difficult in the early stages of the disease and time-consuming. Therefore, multi-task learning is…

图像与视频处理 · 电气工程与系统科学 2024-04-10 Weronika Hryniewska-Guzik , Maria Kędzierska , Przemysław Biecek

Multi-task learning based video anomaly detection methods combine multiple proxy tasks in different branches to detect video anomalies in different situations. Most existing methods either do not combine complementary tasks to effectively…

计算机视觉与模式识别 · 计算机科学 2023-05-12 Mohammad Baradaran , Robert Bergevin

The vulnerability against presentation attacks is a crucial problem undermining the wide-deployment of face recognition systems. Though presentation attack detection (PAD) systems try to address this problem, the lack of generalization and…

计算机视觉与模式识别 · 计算机科学 2022-02-22 Anjith George , David Geissbuhler , Sebastien Marcel

The vast network of bridges in the United States raises a high requirement for maintenance and rehabilitation. The massive cost of manual visual inspection to assess bridge conditions is a burden to some extent. Advanced robots have been…

计算机视觉与模式识别 · 计算机科学 2022-10-31 Chenyu Zhang , Muhammad Monjurul Karim , Ruwen Qin

Basal cell carcinoma (BCC) accounts for about 75% of skin cancers. The adoption of teledermatology protocols in Spanish public hospitals has increased dermatologists' workload, motivating the development of AI tools for lesion…

机器学习 · 计算机科学 2026-03-17 Iván Matas , Carmen Serrano , Francisca Silva , Amalia Serrano , Tomás Toledo-Pastrana , Begoña Acha

In recent years, building change detection methods have made great progress by introducing deep learning, but they still suffer from the problem of the extracted features not being discriminative enough, resulting in incomplete regions and…

计算机视觉与模式识别 · 计算机科学 2019-09-18 Yi Liu , Chao Pang , Zongqian Zhan , Xiaomeng Zhang , Xue Yang

Monocular depth estimation and defocus estimation are two fundamental tasks in computer vision. Most existing methods treat depth estimation and defocus estimation as two separate tasks, ignoring the strong connection between them. In this…

计算机视觉与模式识别 · 计算机科学 2022-08-23 Renzhi He , Hualin Hong , Boya Fu , Fei Liu