English
Related papers

Related papers: MedDChest: A Content-Aware Multimodal Foundational…

200 papers

Deep learning has shown great potential for automated medical image segmentation to improve the precision and speed of disease diagnostics. However, the task presents significant difficulties due to variations in the scale, shape, texture,…

Image and Video Processing · Electrical Eng. & Systems 2024-09-06 Shahzaib Iqbal , Tariq M. Khan , Syed S. Naqvi , Asim Naveed , Erik Meijering

Transformers have demonstrated remarkable performance in natural language processing and computer vision. However, existing vision Transformers struggle to learn from limited medical data and are unable to generalize on diverse medical…

Image and Video Processing · Electrical Eng. & Systems 2023-04-06 Yunhe Gao , Mu Zhou , Di Liu , Zhennan Yan , Shaoting Zhang , Dimitris N. Metaxas

Medical image segmentation is crucial for disease diagnosis and treatment planning, yet developing robust segmentation models often requires substantial computational resources and large datasets. Existing research shows that pre-trained…

Computer Vision and Pattern Recognition · Computer Science 2025-08-05 Paul Zaha , Lars Böcking , Simeon Allmendinger , Leopold Müller , Niklas Kühl

Chest computed tomography (CT) is central to the detection and management of thoracic disease, yet the growing scale and complexity of volumetric imaging increasingly exceed what can be addressed by scan-level prediction alone. Clinically…

Computer Vision and Pattern Recognition · Computer Science 2026-04-28 Xuguang Bai , Mingxuan Liu , Tongxi Song , Yifei Chen , Hongjia Yang , Kasidit Anmahapong , Zihan Li , Ying Zhou , Qiyuan Tian

Transformers have shown great success in medical image segmentation. However, transformers may exhibit a limited generalization ability due to the underlying single-scale self-attention (SA) mechanism. In this paper, we address this issue…

Computer Vision and Pattern Recognition · Computer Science 2023-03-30 Md Mostafijur Rahman , Radu Marculescu

Deep learning-based computer-aided diagnosis is gradually deployed to review and analyze medical images. However, this paradigm is restricted in real-world clinical applications due to the poor robustness and generalization. The issue is…

Computer Vision and Pattern Recognition · Computer Science 2022-03-22 Yurong Chen

Recent advancements in mixed-modal generative have opened new avenues for developing unified biomedical assistants capable of analyzing biomedical images, answering complex questions about them, and generating multimodal patient reports.…

Artificial Intelligence · Computer Science 2025-04-24 Hritik Bansal , Daniel Israel , Siyan Zhao , Shufan Li , Tung Nguyen , Aditya Grover

Clinical decision-making relies on the integration of information across various data modalities, such as clinical time-series, medical images and textual reports. Compared to other domains, real-world medical data is heterogeneous in…

Image and Video Processing · Electrical Eng. & Systems 2025-08-14 Baraa Al Jorf , Farah Shamout

Masked Autoencoders (MAEs) have been shown to be effective in pre-training Vision Transformers (ViTs) for natural and medical image analysis problems. By reconstructing missing pixel/voxel information in visible patches, a ViT encoder can…

Computer Vision and Pattern Recognition · Computer Science 2025-11-20 Pengfei Gu , Huimin Li , Yejia Zhang , Chaoli Wang , Danny Z. Chen

Autoregressive modeling has driven major advances in multimodal AI, yet its application to medical imaging remains constrained by the absence of a unified image tokenizer that simultaneously preserves fine-grained anatomical structures and…

Image and Video Processing · Electrical Eng. & Systems 2026-04-02 Chenglong Ma , Yuanfeng Ji , Jin Ye , Zilong Li , Chenhui Wang , Junzhi Ning , Wei Li , Lihao Liu , Qiushan Guo , Tianbin Li , Junjun He , Hongming Shan

The utilisation of deep learning segmentation algorithms that learn complex organs and tissue patterns and extract essential regions of interest from the noisy background to improve the visual ability for medical image diagnosis has…

Computer Vision and Pattern Recognition · Computer Science 2023-11-03 Yanming Guo

Automated endoscopy video analysis is a challenging task in medical computer vision, with the primary objective of assisting surgeons during procedures. The difficulty arises from the complexity of surgical scenes and the lack of a…

Computer Vision and Pattern Recognition · Computer Science 2023-04-03 Dominik Batić , Felix Holm , Ege Özsoy , Tobias Czempiel , Nassir Navab

Self-supervised learning has greatly facilitated medical image analysis by suppressing the training data requirement for real-world applications. Current paradigms predominantly rely on self-supervision within uni-modal image data, thereby…

Computer Vision and Pattern Recognition · Computer Science 2025-03-31 Shaohao Rui , Lingzhi Chen , Zhenyu Tang , Lilong Wang , Mianxin Liu , Shaoting Zhang , Xiaosong Wang

Medical images are often characterized by their structured anatomical representations and spatially inhomogeneous contrasts. Leveraging anatomical priors in neural networks can greatly enhance their utility in resource-constrained clinical…

Image and Video Processing · Electrical Eng. & Systems 2024-02-07 Xiang Chen , Min Liu , Rongguang Wang , Renjiu Hu , Dongdong Liu , Gaolei Li , Hang Zhang

Large-scale pre-trained models, such as Vision Foundation Models (VFMs), have demonstrated impressive performance across various downstream tasks by transferring generalized knowledge, especially when target data is limited. However, their…

Computer Vision and Pattern Recognition · Computer Science 2025-03-11 Pengchen Liang , Haishan Huang , Bin Pu , Jianguo Chen , Xiang Hua , Jing Zhang , Weibo Ma , Zhuangzhuang Chen , Yiwei Li , Qing Chang

The chest X-ray is often utilized for diagnosing common thoracic diseases. In recent years, many approaches have been proposed to handle the problem of automatic diagnosis based on chest X-rays. However, the scarcity of labeled data for…

Image and Video Processing · Electrical Eng. & Systems 2023-06-05 Weizhi Nie , Chen Zhang , Dan Song , Lina Zhao , Yunpeng Bai , Keliang Xie , Anan Liu

The "pre-training then fine-tuning (FT)" paradigm is widely adopted to boost the model performance of deep learning-based methods for medical volumetric segmentation. However, conventional full FT incurs high computational and memory costs.…

Computer Vision and Pattern Recognition · Computer Science 2024-05-29 Jiachen Shen , Wenxuan Wang , Chen Chen , Jianbo Jiao , Jing Liu , Yan Zhang , Shanshan Song , Jiangyun Li

Prior to the deep learning era, shape was commonly used to describe the objects. Nowadays, state-of-the-art (SOTA) algorithms in medical imaging are predominantly diverging from computer vision, where voxel grids, meshes, point clouds, and…

Computer Vision and Pattern Recognition · Computer Science 2025-06-06 Jianning Li , Zongwei Zhou , Jiancheng Yang , Antonio Pepe , Christina Gsaxner , Gijs Luijten , Chongyu Qu , Tiezheng Zhang , Xiaoxi Chen , Wenxuan Li , Marek Wodzinski , Paul Friedrich , Kangxian Xie , Yuan Jin , Narmada Ambigapathy , Enrico Nasca , Naida Solak , Gian Marco Melito , Viet Duc Vu , Afaque R. Memon , Christopher Schlachta , Sandrine De Ribaupierre , Rajnikant Patel , Roy Eagleson , Xiaojun Chen , Heinrich Mächler , Jan Stefan Kirschke , Ezequiel de la Rosa , Patrick Ferdinand Christ , Hongwei Bran Li , David G. Ellis , Michele R. Aizenberg , Sergios Gatidis , Thomas Küstner , Nadya Shusharina , Nicholas Heller , Vincent Andrearczyk , Adrien Depeursinge , Mathieu Hatt , Anjany Sekuboyina , Maximilian Löffler , Hans Liebl , Reuben Dorent , Tom Vercauteren , Jonathan Shapey , Aaron Kujawa , Stefan Cornelissen , Patrick Langenhuizen , Achraf Ben-Hamadou , Ahmed Rekik , Sergi Pujades , Edmond Boyer , Federico Bolelli , Costantino Grana , Luca Lumetti , Hamidreza Salehi , Jun Ma , Yao Zhang , Ramtin Gharleghi , Susann Beier , Arcot Sowmya , Eduardo A. Garza-Villarreal , Thania Balducci , Diego Angeles-Valdez , Roberto Souza , Leticia Rittner , Richard Frayne , Yuanfeng Ji , Vincenzo Ferrari , Soumick Chatterjee , Florian Dubost , Stefanie Schreiber , Hendrik Mattern , Oliver Speck , Daniel Haehn , Christoph John , Andreas Nürnberger , João Pedrosa , Carlos Ferreira , Guilherme Aresta , António Cunha , Aurélio Campilho , Yannick Suter , Jose Garcia , Alain Lalande , Vicky Vandenbossche , Aline Van Oevelen , Kate Duquesne , Hamza Mekhzoum , Jef Vandemeulebroucke , Emmanuel Audenaert , Claudia Krebs , Timo van Leeuwen , Evie Vereecke , Hauke Heidemeyer , Rainer Röhrig , Frank Hölzle , Vahid Badeli , Kathrin Krieger , Matthias Gunzer , Jianxu Chen , Timo van Meegdenburg , Amin Dada , Miriam Balzer , Jana Fragemann , Frederic Jonske , Moritz Rempe , Stanislav Malorodov , Fin H. Bahnsen , Constantin Seibold , Alexander Jaus , Zdravko Marinov , Paul F. Jaeger , Rainer Stiefelhagen , Ana Sofia Santos , Mariana Lindo , André Ferreira , Victor Alves , Michael Kamp , Amr Abourayya , Felix Nensa , Fabian Hörst , Alexander Brehmer , Lukas Heine , Yannik Hanusrichter , Martin Weßling , Marcel Dudda , Lars E. Podleska , Matthias A. Fink , Julius Keyl , Konstantinos Tserpes , Moon-Sung Kim , Shireen Elhabian , Hans Lamecker , Dženan Zukić , Beatriz Paniagua , Christian Wachinger , Martin Urschler , Luc Duong , Jakob Wasserthal , Peter F. Hoyer , Oliver Basu , Thomas Maal , Max J. H. Witjes , Gregor Schiele , Ti-chiun Chang , Seyed-Ahmad Ahmadi , Ping Luo , Bjoern Menze , Mauricio Reyes , Thomas M. Deserno , Christos Davatzikos , Behrus Puladi , Pascal Fua , Alan L. Yuille , Jens Kleesiek , Jan Egger

Medical image analysis tasks often focus on regions or structures located in a particular location within the patient's body. Often large parts of the image may not be of interest for the image analysis task. When using deep-learning based…

Computer Vision and Pattern Recognition · Computer Science 2024-11-01 Thomas Buddenkotte , Roland Opfer , Julia Krüger , Alessa Hering , Mireia Crispin-Ortuzar

Building AI models with trustworthiness is important especially in regulated areas such as healthcare. In tackling COVID-19, previous work uses convolutional neural networks as the backbone architecture, which has shown to be prone to…

Image and Video Processing · Electrical Eng. & Systems 2022-07-20 Kai Ma , Pengcheng Xi , Karim Habashy , Ashkan Ebadi , Stéphane Tremblay , Alexander Wong