English
Related papers

Related papers: Self-supervised 3D Patient Modeling with Multi-mod…

200 papers

Multimodal fusion has emerged as a promising paradigm for disease diagnosis and prognosis, integrating complementary information from heterogeneous data sources such as medical images, clinical records, and radiology reports. However,…

Computer Vision and Pattern Recognition · Computer Science 2026-02-03 Chongyu Qu , Zhengyi Lu , Yuxiang Lai , Thomas Z. Li , Junchao Zhu , Junlin Guo , Juming Xiong , Yanfan Zhu , Yuechen Yang , Allen J. Luna , Kim L. Sandler , Bennett A. Landman , Yuankai Huo

This paper explores the use of self-supervised deep learning in medical imaging in cases where two scan modalities are available for the same subject. Specifically, we use a large publicly-available dataset of over 20,000 subjects from the…

Computer Vision and Pattern Recognition · Computer Science 2021-08-09 Rhydian Windsor , Amir Jamaludin , Timor Kadir , Andrew Zisserman

The scarcity of well-annotated medical datasets requires leveraging transfer learning from broader datasets like ImageNet or pre-trained models like CLIP. Model soups averages multiple fine-tuned models aiming to improve performance on…

Computer Vision and Pattern Recognition · Computer Science 2024-06-04 Santosh Sanjeev , Nuren Zhaksylyk , Ibrahim Almakky , Anees Ur Rehman Hashmi , Mohammad Areeb Qazi , Mohammad Yaqub

Precise patient positioning is fundamental to successful removal of malignant tumors during treatment of head and neck cancers. Errors in patient positioning have been known to damage critical organs and cause complications. To better…

Robotics · Computer Science 2017-09-26 Olalekan Ogunmolu , Adwait Kulkarni , Yonas Tadesse , Xuejun Gu , Steve Jiang , Nicholas Gans

A key challenge in learning from multimodal biological data is missing modalities, where data from one or more modalities are absent for some patients. Existing approaches either exclude patients with missing modalities, impute missing…

Machine Learning · Computer Science 2026-05-19 Sina Tabakhi , Chen , Chen , Haiping Lu

Although fully autonomous systems still face challenges due to patients' anatomical variability, teleoperated systems appear to be more practical in current healthcare settings. This paper presents an anatomy-aware control framework for…

Robotics · Computer Science 2026-02-13 Davide Nardi , Edoardo Lamon , Daniele Fontanelli , Matteo Saveriano , Luigi Palopoli

Federated learning enables collaborative model training across medical institutions without sharing raw data, but its performance is often limited by domain heterogeneity across clients. Existing approaches to address this challenge fall…

Computer Vision and Pattern Recognition · Computer Science 2026-05-05 Chamani Shiranthika , Parvaneh Saeedi

Multi-organ segmentation is one of most successful applications of deep learning in medical image analysis. Deep convolutional neural nets (CNNs) have shown great promise in achieving clinically applicable image segmentation performance on…

Image and Video Processing · Electrical Eng. & Systems 2020-12-18 Hao Tang , Xingwei Liu , Kun Han , Shanlin Sun , Narisu Bai , Xuming Chen , Huang Qian , Yong Liu , Xiaohui Xie

Monocular 3D human pose estimation poses significant challenges due to the inherent depth ambiguities that arise during the reprojection process from 2D to 3D. Conventional approaches that rely on estimating an over-fit projection matrix…

Computer Vision and Pattern Recognition · Computer Science 2024-01-19 Junkun Jiang , Jie Chen

Motivated by the increasing popularity of attention mechanisms, we observe that popular convolutional (conv.) attention models like Squeeze-and-Excite (SE) and Convolutional Block Attention Module (CBAM) rely on expensive multi-layer…

Computer Vision and Pattern Recognition · Computer Science 2024-07-22 Majedaldein Almahasneh , Xianghua Xie , Adeline Paiement

Modern medicine requires generalised approaches to the synthesis and integration of multimodal data, often at different biological scales, that can be applied to a variety of evidence structures, such as complex disease analyses and…

Quantitative Methods · Quantitative Biology 2019-11-11 Devin Taylor , Simeon Spasov , Pietro Liò

Detection Transformers represent end-to-end object detection approaches based on a Transformer encoder-decoder architecture, exploiting the attention mechanism for global relation modeling. Although Detection Transformers deliver results on…

Computer Vision and Pattern Recognition · Computer Science 2023-06-30 Bastian Wittmann , Fernando Navarro , Suprosanna Shit , Bjoern Menze

Medical image fusion integrates the complementary diagnostic information of the source image modalities for improved visualization and analysis of underlying anomalies. Recently, deep learning-based models have excelled the conventional…

Image and Video Processing · Electrical Eng. & Systems 2023-10-19 Manisha Das , Deep Gupta , Petia Radeva , Ashwini M Bakde

Medical image segmentation - the prerequisite of numerous clinical needs - has been significantly prospered by recent advances in convolutional neural networks (CNNs). However, it exhibits general limitations on modeling explicit long-range…

Computer Vision and Pattern Recognition · Computer Science 2021-07-13 Yundong Zhang , Huiye Liu , Qiang Hu

In the field of medical imaging, AI-assisted techniques such as object detection, segmentation, and classification are widely employed to alleviate the workload of physicians and doctors. However, single-task models are predominantly used,…

Image and Video Processing · Electrical Eng. & Systems 2025-11-18 Fan Li , Arun Iyengar , Lanyu Xu

Manual annotation of large-scale point cloud dataset for varying tasks such as 3D object classification, segmentation and detection is often laborious owing to the irregular structure of point clouds. Self-supervised learning, which…

Computer Vision and Pattern Recognition · Computer Science 2022-03-25 Mohamed Afham , Isuru Dissanayake , Dinithi Dissanayake , Amaya Dharmasiri , Kanchana Thilakarathna , Ranga Rodrigo

Domain adaptation for Cross-LiDAR 3D detection is challenging due to the large gap on the raw data representation with disparate point densities and point arrangements. By exploring domain-invariant 3D geometric characteristics and motion…

Computer Vision and Pattern Recognition · Computer Science 2022-12-02 Xidong Peng , Xinge Zhu , Yuexin Ma

In this survey, we first introduce the background of popular sensors used for self-driving, their data properties, and the corresponding object detection algorithms. Next, we discuss existing datasets that can be used for evaluating…

Computer Vision and Pattern Recognition · Computer Science 2023-03-08 Yingjie Wang , Qiuyu Mao , Hanqi Zhu , Jiajun Deng , Yu Zhang , Jianmin Ji , Houqiang Li , Yanyong Zhang

We propose a robust and accurate method for estimating the 3D poses of two hands in close interaction from a single color image. This is a very challenging problem, as large occlusions and many confusions between the joints may happen.…

Computer Vision and Pattern Recognition · Computer Science 2022-04-20 Shreyas Hampali , Sayan Deb Sarkar , Mahdi Rad , Vincent Lepetit

Early diagnosis of Alzheimer Diagnostics (AD) is a challenging task due to its subtle and complex clinical symptoms. Deep learning-assisted medical diagnosis using image recognition techniques has become an important research topic in this…

Image and Video Processing · Electrical Eng. & Systems 2024-01-26 Yihao Lin , Ximeng Li , Yan Zhang , Jinshan Tang