中文
相关论文

相关论文: ProstaTD: Bridging Surgical Triplet from Classific…

200 篇论文

In recent years, we have seen a significant interest in data-driven deep learning approaches for video anomaly detection, where an algorithm must determine if specific frames of a video contain abnormal behaviors. However, video anomaly…

计算机视觉与模式识别 · 计算机科学 2023-06-05 Armin Danesh Pazho , Ghazal Alinezhad Noghre , Babak Rahimi Ardabili , Christopher Neff , Hamed Tabkhi

Surgical tool detection is a fundamental task for understanding egocentric open surgery videos. However, detecting surgical tools presents significant challenges due to their highly imbalanced class distribution, similar shapes and similar…

计算机视觉与模式识别 · 计算机科学 2024-11-28 Ryo Fujii , Hideo Saito , Hiroki Kajita

Deep learning based visual trackers entail offline pre-training on large volumes of video datasets with accurate bounding box annotations that are labor-expensive to achieve. We present a new framework to facilitate bounding box annotations…

计算机视觉与模式识别 · 计算机科学 2021-08-10 Kenan Dai , Jie Zhao , Lijun Wang , Dong Wang , Jianhua Li , Huchuan Lu , Xuesheng Qian , Xiaoyun Yang

3D image segmentation is one of the most important and ubiquitous problems in medical image processing. It provides detailed quantitative analysis for accurate disease diagnosis, abnormal detection, and classification. Currently deep…

计算机视觉与模式识别 · 计算机科学 2019-06-19 Zhenxi Zhang , Jie Li , Zhusi Zhong , Zhicheng Jiao , Xinbo Gao

Accurate identification and localization of abnormalities from radiology images play an integral part in clinical diagnosis and treatment planning. Building a highly accurate prediction model for these tasks usually requires a large number…

计算机视觉与模式识别 · 计算机科学 2018-06-22 Zhe Li , Chong Wang , Mei Han , Yuan Xue , Wei Wei , Li-Jia Li , Li Fei-Fei

Medical task-oriented dialogue systems can assist doctors by collecting patient medical history, aiding in diagnosis, or guiding treatment selection, thereby reducing doctor burnout and expanding access to medical services. However,…

计算与语言 · 计算机科学 2024-10-21 Vishal Vivek Saley , Goonjan Saha , Rocktim Jyoti Das , Dinesh Raghu , Mausam

Existing 3D pose datasets of object categories are limited to generic object types and lack of fine-grained information. In this work, we introduce a new large-scale dataset that consists of 409 fine-grained categories and 31,881 images…

计算机视觉与模式识别 · 计算机科学 2018-10-23 Yaming Wang , Xiao Tan , Yi Yang , Ziyu Li , Xiao Liu , Feng Zhou , Larry S. Davis

Recent advances in deep learning have transformed computer-assisted intervention and surgical video analysis, driving improvements not only in surgical training, intraoperative decision support, and patient outcomes, but also in…

计算机视觉与模式识别 · 计算机科学 2025-06-16 Sahar Nasirihaghighi , Negin Ghamsarian , Leonie Peschek , Matteo Munari , Heinrich Husslein , Raphael Sznitman , Klaus Schoeffmann

Recent development of object detection mainly depends on deep learning with large-scale benchmarks. However, collecting such fully-annotated data is often difficult or expensive for real-world applications, which restricts the power of deep…

计算机视觉与模式识别 · 计算机科学 2020-02-19 Hao Chen , Yali Wang , Guoyou Wang , Xiang Bai , Yu Qiao

The advancement of machine learning algorithms in medical image analysis requires the expansion of training datasets. A popular and cost-effective approach is automated annotation extraction from free-text medical reports, primarily due to…

计算机视觉与模式识别 · 计算机科学 2025-07-22 Veronika Cheplygina , Cathrine Damgaard , Trine Naja Eriksen , Dovile Juodelyte , Amelia Jiménez-Sánchez

Most of the existing object detection works are based on the bounding box annotation: each object has a precise annotated box. However, for rib fractures, the bounding box annotation is very labor-intensive and time-consuming because…

计算机视觉与模式识别 · 计算机科学 2022-07-06 Zhizhong Chai , Huangjing Lin , Luyang Luo , Pheng-Ann Heng , Hao Chen

Prostate gland segmentation from T2-weighted MRI is a critical yet challenging task in clinical prostate cancer assessment. While deep learning-based methods have significantly advanced automated segmentation, most conventional…

图像与视频处理 · 电气工程与系统科学 2025-06-25 Ahmad Mustafa , Reza Rastegar , Ghassan AlRegib

Training 3D object detectors for autonomous driving has been limited to small datasets due to the effort required to generate annotations. Reducing both task complexity and the amount of task switching done by annotators is key to reducing…

机器学习 · 计算机科学 2018-07-18 Jungwook Lee , Sean Walsh , Ali Harakeh , Steven L. Waslander

We present crowdsourcing as an additional modality to aid radiologists in the diagnosis of lung cancer from clinical chest computed tomography (CT) scans. More specifically, a complete workflow is introduced which can help maximize the…

计算机视觉与模式识别 · 计算机科学 2018-09-19 Saeed Boorboor , Saad Nadeem , Ji Hwan Park , Kevin Baker , Arie Kaufman

Supervised training of object detectors requires well-annotated large-scale datasets, whose production is costly. Therefore, some efforts have been made to obtain annotations in economical ways, such as cloud sourcing. However, datasets…

计算机视觉与模式识别 · 计算机科学 2021-12-08 Jiafeng Mao , Qing Yu , Yoko Yamakata , Kiyoharu Aizawa

Accurate 3D reconstruction of hands and instruments is critical for vision-based analysis of ophthalmic microsurgery, yet progress has been hampered by the lack of realistic, large-scale datasets and reliable annotation tools. In this work,…

计算机视觉与模式识别 · 计算机科学 2025-06-02 Ming Hu , Zhengdi Yu , Feilong Tang , Kaiwen Chen , Yulong Li , Imran Razzak , Junjun He , Tolga Birdal , Kaijing Zhou , Zongyuan Ge

There is a dire need for medical imaging datasets with accompanying annotations to perform downstream patient analysis. However, it is difficult to manually generate these annotations, due to the time-consuming nature, and the variability…

图像与视频处理 · 电气工程与系统科学 2024-06-21 Deepa Krishnaswamy , Vamsi Krishna Thiriveedhi , Cosmin Ciausu , David Clunie , Steve Pieper , Ron Kikinis , Andrey Fedorov

Video anomaly retrieval aims to localize anomalous events in videos using natural language queries to facilitate public safety. However, existing datasets suffer from severe limitations: (1) data scarcity due to the long-tail nature of…

计算机视觉与模式识别 · 计算机科学 2025-06-03 Shuyu Yang , Yilun Wang , Yaxiong Wang , Li Zhu , Zhedong Zheng

Image/video coding has been a remarkable research area for both academia and industry for many years. Testing datasets, especially high-quality image/video datasets are desirable for the justified evaluation of coding-related research,…

图像与视频处理 · 电气工程与系统科学 2025-03-18 Zhuoyuan Li , Junqi Liao , Chuanbo Tang , Haotian Zhang , Yuqi Li , Yifan Bian , Xihua Sheng , Xinmin Feng , Yao Li , Changsheng Gao , Li Li , Dong Liu , Feng Wu