中文
相关论文

相关论文: Single- and Multi-Task Architectures for Tool Pres…

200 篇论文

Object detection, segmentation and classification are three common tasks in medical image analysis. Multi-task deep learning (MTL) tackles these three tasks jointly, which provides several advantages saving computing time and resources and…

计算机视觉与模式识别 · 计算机科学 2019-06-06 Fei Gao , Hyunsoo Yoon , Teresa Wu , Xianghua Chu

Advancements in attention mechanisms have led to significant performance improvements in a variety of areas in machine learning due to its ability to enable the dynamic modeling of temporal sequences. A particular area in computer vision…

计算机视觉与模式识别 · 计算机科学 2021-12-14 Brennan Gebotys , Alexander Wong , David A. Clausi

Robotic-assisted procedures offer numerous advantages over traditional approaches, including improved dexterity, reduced fatigue, minimized trauma, and superior outcomes. However, the main challenge of these systems remains the poor…

机器人学 · 计算机科学 2025-01-08 Gabriela Rus , Nadim Al Hajjar , Ionut Zima , Calin Vaida , Corina Radu , Damien Chablat , Andra Ciocan , Doina Pîslă

The recent surge in performance for image analysis of digitised pathology slides can largely be attributed to the advances in deep learning. Deep models can be used to initially localise various structures in the tissue and hence facilitate…

图像与视频处理 · 电气工程与系统科学 2022-11-15 Simon Graham , Quoc Dang Vu , Mostafa Jahanifar , Shan E Ahmed Raza , Fayyaz Minhas , David Snead , Nasir Rajpoot

Automatic recognition of fine-grained surgical activities, called steps, is a challenging but crucial task for intelligent intra-operative computer assistance. The development of current vision-based activity recognition methods relies…

计算机视觉与模式识别 · 计算机科学 2023-04-12 Sanat Ramesh , Diego Dall'Alba , Cristians Gonzalez , Tong Yu , Pietro Mascagni , Didier Mutter , Jacques Marescaux , Paolo Fiorini , Nicolas Padoy

Surgical phase recognition is a basic component for different context-aware applications in computer- and robot-assisted surgery. In recent years, several methods for automatic surgical phase recognition have been proposed, showing…

计算机视觉与模式识别 · 计算机科学 2023-05-24 Isabel Funke , Dominik Rivoir , Stefanie Speidel

Many objects do not appear frequently enough in complex scenes (e.g., certain handbags in living rooms) for training an accurate object detector, but are often found frequently by themselves (e.g., in product images). Yet, these…

计算机视觉与模式识别 · 计算机科学 2021-09-14 Cheng Zhang , Tai-Yu Pan , Yandong Li , Hexiang Hu , Dong Xuan , Soravit Changpinyo , Boqing Gong , Wei-Lun Chao

Despite the recent progress in video understanding and the continuous rate of improvement in temporal action localization throughout the years, it is still unclear how far (or close?) we are to solving the problem. To this end, we introduce…

计算机视觉与模式识别 · 计算机科学 2018-07-30 Humam Alwassel , Fabian Caba Heilbron , Victor Escorcia , Bernard Ghanem

Medical image denoising is considered among the most challenging vision tasks. Despite the real-world implications, existing denoising methods have notable drawbacks as they often generate visual artifacts when applied to heterogeneous…

图像与视频处理 · 电气工程与系统科学 2025-03-11 S M A Sharif , Rizwan Ali Naqvi , Woong-Kee Loh

In this paper, we present a two-stream multi-task network for fashion recognition. This task is challenging as fashion clothing always contain multiple attributes, which need to be predicted simultaneously for real-time industrial systems.…

计算机视觉与模式识别 · 计算机科学 2019-05-14 Peizhao Li , Yanjing Li , Xiaolong Jiang , Xiantong Zhen

Following recent advancements in computer-aided detection and diagnosis systems for colonoscopy, the automated reporting of colonoscopy procedures is set to further revolutionize clinical practice. A crucial yet underexplored aspect in the…

计算机视觉与模式识别 · 计算机科学 2025-04-28 Carlo Biffi , Giorgio Roffo , Pietro Salvagnini , Andrea Cherubini

Multi-task visual anomaly detection is critical for car-related manufacturing quality assessment. However, existing methods remain task-specific, hindered by the absence of a unified benchmark for multi-task evaluation. To fill in this gap,…

计算机视觉与模式识别 · 计算机科学 2026-04-13 Jiahua Pang , Ying Li , Dongpu Cao , Jingcai Luo , Yanuo Zheng , Bao Yunfan , Yujie Lei , Rui Yuan , Yuxi Tian , Guojin Yuan , Hongchang Chen , Zhi Zheng , Yongchun Liu

The field of computer vision applied to videos of minimally invasive surgery is ever-growing. Workflow recognition pertains to the automated recognition of various aspects of a surgery: including which surgical steps are performed; and…

We propose a novel convolutional neural network approach to address the fine-grained recognition problem of multi-view dynamic facial action unit detection. We leverage recent gains in large-scale object recognition by formulating the task…

计算机视觉与模式识别 · 计算机科学 2018-08-21 Andres Romero , Juan Leon , Pablo Arbelaez

Understanding the content of videos is one of the core techniques for developing various helpful applications in the real world, such as recognizing various human actions for surveillance systems or customer behavior analysis in an…

计算机视觉与模式识别 · 计算机科学 2019-07-12 Chiwan Song , Woobin Im , Sung-eui Yoon

Clinical cystoscopy, the current standard for bladder cancer diagnosis, suffers from significant reliance on physician expertise, leading to variability and subjectivity in diagnostic outcomes. There is an urgent need for objective,…

图像与视频处理 · 电气工程与系统科学 2025-08-22 Jinliang Yu , Mingduo Xie , Yue Wang , Tianfan Fu , Xianglai Xu , Jiajun Wang

Depth completion and object detection are two crucial tasks often used for aerial 3D mapping, path planning, and collision avoidance of Uncrewed Aerial Vehicles (UAVs). Common solutions include using measurements from a LiDAR sensor;…

计算机视觉与模式识别 · 计算机科学 2023-04-26 Sara Hatami Gazani , Fardad Dadboud , Miodrag Bolic , Iraj Mantegh , Homayoun Najjaran

Computer-Assisted Intervention (CAI) has the potential to revolutionize modern surgery, with surgical scene understanding serving as a critical component in supporting decision-making, improving procedural efficacy, and ensuring…

Laparoscopic video tracking primarily focuses on two target types: surgical instruments and anatomy. The former could be used for skill assessment, while the latter is necessary for the projection of virtual overlays. Where instrument and…

计算机视觉与模式识别 · 计算机科学 2024-03-29 Beerend G. A. Gerats , Jelmer M. Wolterink , Seb P. Mol , Ivo A. M. J. Broeders

Multi-task learning is commonly used in autonomous driving for solving various visual perception tasks. It offers significant benefits in terms of both performance and computational complexity. Current work on multi-task learning networks…

计算机视觉与模式识别 · 计算机科学 2019-04-23 Sumanth Chennupati , Ganesh Sistu , Senthil Yogamani , Samir A Rawashdeh