中文
相关论文

相关论文: A Multimodal Approach For Endoscopic VCE Image Cla…

200 篇论文

Automated radiology report generation aims to expedite the tedious and error-prone reporting process for radiologists. While recent works have made progress, learning to align medical images and textual findings remains challenging due to…

计算机视觉与模式识别 · 计算机科学 2025-03-21 Yaxiong Chen , Chuang Du , Chunlei Li , Jingliang Hu , Yilei Shi , Shengwu Xiong , Xiao Xiang Zhu , Lichao Mou

Medical image classification plays a crucial role in computer-aided clinical diagnosis. While deep learning techniques have significantly enhanced efficiency and reduced costs, the privacy-sensitive nature of medical imaging data…

计算机视觉与模式识别 · 计算机科学 2024-07-04 Sufen Ren , Yule Hu , Shengchao Chen , Guanjun Wang

This study evaluated the effect of BioBERT in medical text processing for the task of medical named entity recognition. Through comparative experiments with models such as BERT, ClinicalBERT, SciBERT, and BlueBERT, the results showed that…

计算与语言 · 计算机科学 2024-12-12 Jiacheng Hu , Runyuan Bao , Yang Lin , Hanchao Zhang , Yanlin Xiang

In this paper, we present Endo-SemiS, a semi-supervised segmentation framework for providing reliable segmentation of endoscopic video frames with limited annotation. EndoSemiS uses 4 strategies to improve performance by effectively…

计算机视觉与模式识别 · 计算机科学 2025-12-22 Hao Li , Daiwei Lu , Xing Yao , Nicholas Kavoussi , Ipek Oguz

An innovative few-shot anomaly detection approach is presented, leveraging the pre-trained CLIP model for medical data, and adapting it for both image-level anomaly classification (AC) and pixel-level anomaly segmentation (AS). A…

计算机视觉与模式识别 · 计算机科学 2025-07-01 Mahshid Shiri , Cigdem Beyan , Vittorio Murino

The early detection of a pulmonary embolism (PE) is critical for enhancing patient survival rates. Both image-based and non-image-based features are of utmost importance in medical classification tasks. In a clinical setting, physicians…

图像与视频处理 · 电气工程与系统科学 2024-04-18 Zhaoxin Guo , Zhipeng Wang , Ruiquan Ge , Jianxun Yu , Feiwei Qin , Yuan Tian , Yuqing Peng , Yonghong Li , Changmiao Wang

The recent surge of foundation models in computer vision and natural language processing opens up perspectives in utilizing multi-modal clinical data to train large models with strong generalizability. Yet pathological image datasets often…

计算机视觉与模式识别 · 计算机科学 2023-07-28 Yunkun Zhang , Jin Gao , Mu Zhou , Xiaosong Wang , Yu Qiao , Shaoting Zhang , Dequan Wang

In this paper, we present the VMSE U-Net and VM-Unet CBAM+ model, two cutting-edge deep learning architectures designed to enhance medical image segmentation. Our approach integrates Squeeze-and-Excitation (SE) and Convolutional Block…

图像与视频处理 · 电气工程与系统科学 2025-07-10 Sayandeep Kanrar , Raja Piyush , Qaiser Razi , Debanshi Chakraborty , Vikas Hassija , GSS Chalapathi

Accurately segmenting different organs from medical images is a critical prerequisite for computer-assisted diagnosis and intervention planning. This study proposes a deep learning-based approach for segmenting various organs from CT and…

Early detection of cervical cancer is crucial for improving patient outcomes and reducing mortality by identifying precancerous lesions as soon as possible. As a result, the use of pap smear screening has significantly increased, leading to…

图像与视频处理 · 电气工程与系统科学 2025-10-21 Theo Di Piazza , Loic Boussel

Video endoscopy represents a major advance in the investigation of gastrointestinal diseases. Reviewing endoscopy videos often involves frequent adjustments and reorientations to piece together a complete view, which can be both…

计算机视觉与模式识别 · 计算机科学 2025-02-14 Juming Xiong , Muyang Li , Ruining Deng , Tianyuan Yao , Shunxing Bao , Regina N Tyree , Girish Hiremath , Yuankai Huo

Ambiguity poses persistent challenges in natural language understanding for large language models (LLMs). To better understand how lexical ambiguity can be resolved through the visual domain, we develop an interpretable Visual Word Sense…

计算与语言 · 计算机科学 2026-02-09 Shamik Bhattacharya , Daniel Perkins , Yaren Dogan , Vineeth Konjeti , Sudarshan Srinivasan , Edmon Begoli

Medical image captioning is a challenging task that requires generating clinically accurate and semantically meaningful descriptions of radiology images. While recent vision-language models (VLMs) such as BLIP, BLIP2, Gemini and ViT-GPT2…

图像与视频处理 · 电气工程与系统科学 2025-05-22 Manshi Limbu , Diwita Banerjee

A deep learning-based monocular depth estimation (MDE) technique is proposed for selection of most informative frames (key frames) of an endoscopic video. In most of the cases, ground truth depth maps of polyps are not readily available and…

计算机视觉与模式识别 · 计算机科学 2021-07-12 Pradipta Sasmal , Avinash Paul , M. K. Bhuyan , Yuji Iwahori

This paper presents a comprehensive comparative model analysis on a novel gastrointestinal medical imaging dataset, comprised of 4,000 endoscopic images spanning four critical disease classes: Diverticulosis, Neoplasm, Peritonitis, and…

计算机视觉与模式识别 · 计算机科学 2025-12-01 Walid Houmaidi , Mohamed Hadadi , Youssef Sabiri , Yousra Chtouki

In-vivo optical microscopy is advancing into routine clinical practice for non-invasively guiding diagnosis and treatment of cancer and other diseases, and thus beginning to reduce the need for traditional biopsy. However, reading and…

图像与视频处理 · 电气工程与系统科学 2020-01-07 Kivanc Kose , Alican Bozkurt , Christi Alessi-Fox , Melissa Gill , Caterina Longo , Giovanni Pellacani , Jennifer Dy , Dana H. Brooks , Milind Rajadhyaksha

Data is one of the essential ingredients to power deep learning research. Small datasets, especially specific to medical institutes, bring challenges to deep learning training stage. This work aims to develop a practical deep multimodal…

机器学习 · 计算机科学 2019-02-26 Faik Aydin , Maggie Zhang , Michelle Ananda-Rajah , Gholamreza Haffari

The quality of images captured by wireless capsule endoscopy (WCE) is key for doctors to diagnose diseases of gastrointestinal (GI) tract. However, there exist many low-quality endoscopic images due to the limited illumination and complex…

图像与视频处理 · 电气工程与系统科学 2020-03-17 Shaofeng Zou , Mingzhu Long , Xuyang Wang , Xiang Xie , Guolin Li , Zhihua Wang

Liver cancer is one of the most common malignant diseases in the world. Segmentation and labeling of liver tumors and blood vessels in CT images can provide convenience for doctors in liver tumor diagnosis and surgical intervention. In the…

图像与视频处理 · 电气工程与系统科学 2022-03-01 Xiangyu Meng , Xudong Zhang , Gan Wang , Ying Zhang , Xin Shi , Huanhuan Dai , Zixuan Wang , Xun Wang

Purpose: To design multi-disease classifiers for body CT scans for three different organ systems using automatically extracted labels from radiology text reports.Materials & Methods: This retrospective study included a total of 12,092…

计算机视觉与模式识别 · 计算机科学 2021-12-07 Fakrul Islam Tushar , Vincent M. D'Anniballe , Rui Hou , Maciej A. Mazurowski , Wanyi Fu , Ehsan Samei , Geoffrey D. Rubin , Joseph Y. Lo