English
Related papers

Related papers: Mono-Modalizing Extremely Heterogeneous Multi-Moda…

200 papers

General vision encoders like DINOv2 and SAM have recently transformed computer vision. Even though they are trained on natural images, such encoder models have excelled in medical imaging, e.g., in classification, segmentation, and…

Computer Vision and Pattern Recognition · Computer Science 2024-07-19 Fryderyk Kögl , Anna Reithmeir , Vasiliki Sideri-Lampretsa , Ines Machado , Rickmer Braren , Daniel Rückert , Julia A. Schnabel , Veronika A. Zimmer

Medical Image Retrieval (MIR) helps doctors quickly find similar patients' data, which can considerably aid the diagnosis process. MIR is becoming increasingly helpful due to the wide use of digital imaging modalities and the growth of the…

Computer Vision and Pattern Recognition · Computer Science 2020-07-20 Yang Feng , Yubao Liu , Jiebo Luo

Multimodal medical image fusion plays an instrumental role in several areas of medical image processing, particularly in disease recognition and tumor detection. Traditional fusion methods tend to process each modality independently before…

Image and Video Processing · Electrical Eng. & Systems 2023-10-11 Lin Liu , Xinxin Fan , Chulong Zhang , Jingjing Dai , Yaoqin Xie , Xiaokun Liang

Multimodal pathological images are usually in clinical diagnosis, but computer vision-based multimodal image-assisted diagnosis faces challenges with modality fusion, especially in the absence of expert-annotated data. To achieve the…

Computer Vision and Pattern Recognition · Computer Science 2025-09-24 Qinghua Lin , Guang-Hai Liu , Zuoyong Li , Yang Li , Yuting Jiang , Xiang Wu

Tracking microsctructural changes in the developing brain relies on accurate inter-subject image registration. However, most methods rely on either structural or diffusion data to learn the spatial correspondences between two or more…

Image and Video Processing · Electrical Eng. & Systems 2020-05-15 Irina Grigorescu , Alena Uus , Daan Christiaens , Lucilio Cordero-Grande , Jana Hutter , A. David Edwards , Joseph V. Hajnal , Marc Modat , Maria Deprez

Deformable registration of magnetic resonance images between patients with brain tumors and healthy subjects has been an important tool to specify tumor geometry through location alignment and facilitate pathological analysis. Since tumor…

Image and Video Processing · Electrical Eng. & Systems 2021-01-19 Xiaofeng Liu , Fangxu Xing , Chao Yang , C. -C. Jay Kuo , Georges ElFakhri , Jonghye Woo

In this paper, we present our submission to the LUMIR25 task of Learn2Reg 2025, which ranked 1st overall on the test set. Extended from LUMIR24, this year's task focuses on zero-shot registration under domain shifts (e.g., high-field MRI,…

Image and Video Processing · Electrical Eng. & Systems 2026-02-24 Hengjie Liu , Yimeng Dou , Di Xu , Xinyi Fu , Dan Ruan , Ke Sheng

The use of Augmented Reality (AR) devices for surgical guidance has gained increasing traction in the medical field. Traditional registration methods often rely on external fiducial markers to achieve high accuracy and real-time…

Computer Vision and Pattern Recognition · Computer Science 2025-06-03 Yue Yang , Christoph Leuze , Brian Hargreaves , Bruce Daniel , Fred Baik

This paper presents NimbleReg, a light-weight deep-learning (DL) framework for diffeomorphic image registration leveraging surface representation of multiple segmented anatomical regions. Deep learning has revolutionized image registration…

Computer Vision and Pattern Recognition · Computer Science 2026-04-29 Antoine Legouhy , Ross Callaghan , Nolah Mazet , Vivien Julienne , Hojjat Azadbakht , Hui Zhang

We propose LiftReg, a 2D/3D deformable registration approach. LiftReg is a deep registration framework which is trained using sets of digitally reconstructed radiographs (DRR) and computed tomography (CT) image pairs. By using simulated…

Image and Video Processing · Electrical Eng. & Systems 2023-04-05 Lin Tian , Yueh Z. Lee , Raúl San José Estépar , Marc Niethammer

Multi-modal entity alignment aims to identify equivalent entities between two multi-modal Knowledge graphs by integrating multi-modal data, such as images and text, to enrich the semantic representations of entities. However, existing…

Artificial Intelligence · Computer Science 2026-01-21 Zhifei Li , Ziyue Qin , Xiangyu Luo , Xiaoju Hou , Yue Zhao , Miao Zhang , Zhifang Huang , Kui Xiao , Bing Yang

Multimodal learning leverages complementary information derived from different modalities, thereby enhancing performance in medical image segmentation. However, prevailing multimodal learning methods heavily rely on extensive well-annotated…

Computer Vision and Pattern Recognition · Computer Science 2024-09-05 Xiaogen Zhou , Yiyou Sun , Min Deng , Winnie Chiu Wing Chu , Qi Dou

Multi-spectral optoacoustic tomography (MSOT) is an emerging optical imaging method providing multiplex molecular and functional information from the rodent brain. It can be greatly augmented by magnetic resonance imaging (MRI) that offers…

Image and Video Processing · Electrical Eng. & Systems 2021-09-07 Yexing Hu , Berkan Lafci , Artur Luzgin , Hao Wang , Jan Klohs , Xose Luis Dean-Ben , Ruiqing Ni , Daniel Razansky , Wuwei Ren

Medical image synthesis remains challenging due to misalignment noise during training. Existing methods have attempted to address this challenge by incorporating a registration-guided module. However, these methods tend to overlook the…

Computer Vision and Pattern Recognition · Computer Science 2024-07-11 Chuanpu Li , Zeli Chen , Yiwen Zhang , Liming Zhong , Wei Yang

Foundational models are trained on extensive datasets to capture the general trends of a domain. However, in medical imaging, the scarcity of data makes pre-training for every domain, modality, or task challenging. Instead of building…

Computer Vision and Pattern Recognition · Computer Science 2025-11-17 Mohammad Areeb Qazi , Munachiso S Nwadike , Ibrahim Almakky , Mohammad Yaqub , Numan Saeed

Deformable image registration aims to precisely align medical images from different modalities or times. Traditional deep learning methods, while effective, often lack interpretability, real-time observability and adjustment capacity during…

Computer Vision and Pattern Recognition · Computer Science 2024-10-08 Yongtai Zhuo , Yiqing Shen

In this paper, we summarize the methods and experimental results we proposed for Task 2 in the learn2reg 2024 Challenge. This task focuses on unsupervised registration of anatomical structures in brain MRI images between different patients.…

Computer Vision and Pattern Recognition · Computer Science 2024-09-05 Yuxi Zhang , Xiang Chen , Jiazheng Wang , Min Liu , Yaonan Wang , Dongdong Liu , Renjiu Hu , Hang Zhang

Histopathology and transcriptomics are fundamental modalities in oncology, encapsulating the morphological and molecular aspects of the disease. Multi-modal self-supervised learning has demonstrated remarkable potential in learning…

Computer Vision and Pattern Recognition · Computer Science 2025-11-18 Tianyi Wang , Jianan Fan , Dingxin Zhang , Dongnan Liu , Yong Xia , Heng Huang , Weidong Cai

Co-registration of multimodal remote sensing images is still an ongoing challenge because of nonlinear radiometric differences (NRD) and significant geometric distortions (e.g., scale and rotation changes) between these images. In this…

Computer Vision and Pattern Recognition · Computer Science 2022-03-01 Yuanxin Ye , Bai Zhu , Tengfeng Tang , Chao Yang , Qizhi Xu , Guo Zhang

This work studies the problem of unsupervised RGB-D point cloud registration, which aims at training a robust registration model without ground-truth pose supervision. Existing methods usually leverages unposed RGB-D sequences and adopt a…

Computer Vision and Pattern Recognition · Computer Science 2025-05-02 Zhinan Yu , Zheng Qin , Yijie Tang , Yongjun Wang , Renjiao Yi , Chenyang Zhu , Kai Xu