English
Related papers

Related papers: Enhancing Chest X-ray Classification through Knowl…

200 papers

Contrastive learning has been proved to be a promising technique for image-level representation learning from unlabeled data. Many existing works have demonstrated improved results by applying contrastive learning in classification and…

Image and Video Processing · Electrical Eng. & Systems 2021-09-20 Dewen Zeng , John N. Kheir , Peng Zeng , Yiyu Shi

Chest X-Ray (CXR) classification in clinical practice is often limited by imperfect supervision, arising from (i) extreme long-tailed multi-label disease distributions and (ii) missing annotations for rare or previously unseen findings. The…

Computer Vision and Pattern Recognition · Computer Science 2026-02-24 Ha-Hieu Pham , Hai-Dang Nguyen , Thanh-Huy Nguyen , Min Xu , Ulas Bagci , Trung-Nghia Le , Huy-Hieu Pham

Self-supervised pre-training of deep learning models with contrastive learning is a widely used technique in image analysis. Current findings indicate a strong potential for contrastive pre-training on medical images. However, further…

Image and Video Processing · Electrical Eng. & Systems 2024-10-21 Daniel Wolf , Tristan Payer , Catharina Silvia Lisson , Christoph Gerhard Lisson , Meinrad Beer , Michael Götz , Timo Ropinski

Although deep learning models for chest X-ray interpretation are commonly trained on labels generated by automatic radiology report labelers, the impact of improvements in report labeling on the performance of chest X-ray classification…

Image and Video Processing · Electrical Eng. & Systems 2021-11-30 Saahil Jain , Akshay Smit , Andrew Y. Ng , Pranav Rajpurkar

In medical image classification tasks, it is common to find that the number of normal samples far exceeds the number of abnormal samples. In such class-imbalanced situations, reliable training of deep neural networks continues to be a major…

Machine Learning · Computer Science 2022-04-06 Sivaramakrishnan Rajaraman , Prasanth Ganesan , Sameer Antani

We present XKD, a novel self-supervised framework to learn meaningful representations from unlabelled videos. XKD is trained with two pseudo objectives. First, masked data reconstruction is performed to learn modality-specific…

Computer Vision and Pattern Recognition · Computer Science 2023-12-27 Pritam Sarkar , Ali Etemad

While multi-modal foundation models pre-trained on large-scale data have been successful in natural language understanding and vision recognition, their use in medical domains is still limited due to the fine-grained nature of medical tasks…

Computer Vision and Pattern Recognition · Computer Science 2023-06-16 Xiaoman Zhang , Chaoyi Wu , Ya Zhang , Yanfeng Wang , Weidi Xie

Medical image analysis often faces significant challenges due to limited expert-annotated data, hindering both model generalization and clinical adoption. We propose an expert-guided explainable few-shot learning framework that integrates…

Image and Video Processing · Electrical Eng. & Systems 2025-09-12 Ifrat Ikhtear Uddin , Longwei Wang , KC Santosh

Multimodal deep learning utilizing imaging and diagnostic reports has made impressive progress in the field of medical imaging diagnostics, demonstrating a particularly strong capability for auxiliary diagnosis in cases where sufficient…

Computer Vision and Pattern Recognition · Computer Science 2024-08-21 Hao Yang , Hong-Yu Zhou , Cheng Li , Weijian Huang , Jiarun Liu , Yong Liang , Guangming Shi , Hairong Zheng , Qiegen Liu , Shanshan Wang

The large-scale development of large language models (LLMs) in medical contexts, such as diagnostic assistance and treatment recommendations, necessitates that these models possess accurate medical knowledge and deliver traceable…

Artificial Intelligence · Computer Science 2025-08-12 Qiyuan Li , Haijiang Liu , Caicai Guo , Chao Gao , Deyu Chen , Meng Wang , Feng Gao , Frank van Harmelen , Jinguang Gu

Medical imaging plays a vital role in modern diagnostics; however, interpreting high-resolution radiological data remains time-consuming and susceptible to variability among clinicians. Traditional image processing techniques often lack the…

Computer Vision and Pattern Recognition · Computer Science 2025-10-21 Melika Filvantorkaman , Maral Filvan Torkaman

Semi-supervised learning addresses the issue of limited annotations in medical images effectively, but its performance is often inadequate for complex backgrounds and challenging tasks. Multi-modal fusion methods can significantly improve…

Computer Vision and Pattern Recognition · Computer Science 2025-06-23 Dongdong Meng , Sheng Li , Hao Wu , Guoping Wang , Xueqing Yan

Medical vision-language pretraining models (VLPM) have achieved remarkable progress in fusing chest X-rays (CXR) with clinical texts, introducing image-text data binding approaches that enable zero-shot learning and downstream clinical…

Computer Vision and Pattern Recognition · Computer Science 2024-03-21 Yuan Gao , Sangwook Kim , David E Austin , Chris McIntosh

The chest X-ray is often utilized for diagnosing common thoracic diseases. In recent years, many approaches have been proposed to handle the problem of automatic diagnosis based on chest X-rays. However, the scarcity of labeled data for…

Image and Video Processing · Electrical Eng. & Systems 2023-06-05 Weizhi Nie , Chen Zhang , Dan Song , Lina Zhao , Yunpeng Bai , Keliang Xie , Anan Liu

The chest X-ray (CXR) is commonly employed to diagnose thoracic illnesses, but the challenge of achieving accurate automatic diagnosis through this method persists due to the complex relationship between pathology. In recent years, various…

Image and Video Processing · Electrical Eng. & Systems 2023-05-23 Weizhi Nie , Chen Zhang , Dan song , Yunpeng Bai , Keliang Xie , Anan Liu

An important component of human analysis of medical images and their context is the ability to relate newly seen things to related instances in our memory. In this paper we mimic this ability by using multi-modal retrieval augmentation and…

Computer Vision and Pattern Recognition · Computer Science 2023-02-23 Tom van Sonsbeek , Marcel Worring

In the domain of vision-language integration, generating detailed image captions poses a significant challenge due to the lack of curated and rich datasets. This study introduces PixLore, a novel method that leverages Querying Transformers…

Recent advances in multimodal ECG representation learning center on aligning ECG signals with paired free-text reports. However, suboptimal alignment persists due to the complexity of medical language and the reliance on a full 12-lead…

Machine Learning · Computer Science 2025-02-26 Che Liu , Cheng Ouyang , Zhongwei Wan , Haozhe Wang , Wenjia Bai , Rossella Arcucci

In recent times, the use of chest Computed Tomography (CT) images for detecting coronavirus infections has gained significant attention, owing to their ability to reveal bilateral changes in affected individuals. However, classifying…

Image and Video Processing · Electrical Eng. & Systems 2023-10-27 Amir Ali

Radiologists highly desire fully automated versatile AI for medical imaging interpretation. However, the lack of extensively annotated large-scale multi-disease datasets has hindered the achievement of this goal. In this paper, we explore…

Computer Vision and Pattern Recognition · Computer Science 2024-04-09 Weiwei Cao , Jianpeng Zhang , Yingda Xia , Tony C. W. Mok , Zi Li , Xianghua Ye , Le Lu , Jian Zheng , Yuxing Tang , Ling Zhang
‹ Prev 1 8 9 10 Next ›