English
Related papers

Related papers: Anatomical Attention Alignment representation for …

200 papers

For robot-assisted surgery, an accurate surgical report reflects clinical operations during surgery and helps document entry tasks, post-operative analysis and follow-up treatment. It is a challenging task due to many complex and diverse…

Computer Vision and Pattern Recognition · Computer Science 2023-08-30 Hongqiu Wang , Yueming Jin , Lei Zhu

The key to facial expression recognition is to learn discriminative spatial-temporal representations that embed facial expression dynamics. Previous studies predominantly rely on pre-trained Convolutional Neural Networks (CNNs) to learn…

Computer Vision and Pattern Recognition · Computer Science 2025-12-01 Yan Li , Yong Zhao , Xiaohan Xia , Dongmei Jiang

Graph neural network (GNN) models are increasingly being used for the classification of electroencephalography (EEG) data. However, GNN-based diagnosis of neurological disorders, such as Alzheimer's disease (AD), remains a relatively…

Neurons and Cognition · Quantitative Biology 2023-12-21 Dominik Klepl , Fei He , Min Wu , Daniel J. Blackburn , Ptolemaios G. Sarrigiannis

Multimodal models trained on large natural image-text pair datasets have exhibited astounding abilities in generating high-quality images. Medical imaging data is fundamentally different to natural images, and the language used to…

Convolutional neural networks have been widely applied to medical image segmentation and have achieved considerable performance. However, the performance may be significantly affected by the domain gap between training data (source domain)…

Image and Video Processing · Electrical Eng. & Systems 2022-07-28 Junyan Lyu , Yiqi Zhang , Yijin Huang , Li Lin , Pujin Cheng , Xiaoying Tang

Different types of staining highlight different structures in organs, thereby assisting in diagnosis. However, due to the impossibility of repeated staining, we cannot obtain different types of stained slides of the same tissue area.…

Image and Video Processing · Electrical Eng. & Systems 2024-04-17 Zexin Li , Yiyang Lin , Zijie Fang , Shuyan Li , Xiu Li

Accurate three-dimensional (3D) tooth segmentation from Cone-Beam Computed Tomography (CBCT) is a prerequisite for digital dental workflows. However, achieving high-fidelity segmentation remains challenging due to adhesion artifacts in…

Computer Vision and Pattern Recognition · Computer Science 2026-01-13 Bing Yu , Liu Shi , Haitao Wang , Deran Qi , Xiang Cai , Wei Zhong , Qiegen Liu

Conversational AI tools that can generate and discuss clinically correct radiology reports for a given medical image have the potential to transform radiology. Such a human-in-the-loop radiology assistant could facilitate a collaborative…

Computer Vision and Pattern Recognition · Computer Science 2025-05-08 Chantal Pellegrini , Ege Özsoy , Benjamin Busam , Nassir Navab , Matthias Keicher

Large language models (LLMs) often generate outdated or inaccurate information based on static training datasets. Retrieval-augmented generation (RAG) mitigates this by integrating outside data sources. While previous RAG systems used…

Generative models have revolutionized Artificial Intelligence (AI), particularly in multimodal applications. However, adapting these models to the medical domain poses unique challenges due to the complexity of medical data and the…

Computer Vision and Pattern Recognition · Computer Science 2026-02-16 Daniele Molino , Francesco di Feola , Linlin Shen , Paolo Soda , Valerio Guarrasi

Our paper focuses on automating the generation of medical reports from chest X-ray image inputs, a critical yet time-consuming task for radiologists. Unlike existing medical re-port generation efforts that tend to produce human-readable…

Computation and Language · Computer Science 2021-08-30 Hoang T. N. Nguyen , Dong Nie , Taivanbat Badamdorj , Yujie Liu , Yingying Zhu , Jason Truong , Li Cheng

Multimodal MR image synthesis aims to generate missing modality images by effectively fusing and mapping from a subset of available MRI modalities. Most existing methods adopt an image-to-image translation paradigm, treating multiple…

Image and Video Processing · Electrical Eng. & Systems 2025-04-29 Tao Song , Yicheng Wu , Minhao Hu , Xiangde Luo , Linda Wei , Guotai Wang , Yi Guo , Feng Xu , Shaoting Zhang

Prediction the conversion to early-stage dementia is critical for mitigating its progression but remains challenging due to subtle cognitive impairments and structural brain changes. Traditional T1-weighted magnetic resonance imaging…

Computer Vision and Pattern Recognition · Computer Science 2024-06-28 Yilin Leng , Wenju Cui , Bai Chen , Xi Jiang , Shuangqing Chen , Jian Zheng

Predicting medications is a crucial task in many intelligent healthcare systems. It can assist doctors in making informed medication decisions for patients according to electronic medical records (EMRs). However, medication prediction is a…

Artificial Intelligence · Computer Science 2022-05-02 Yang An , Bo Jin , Xiaopeng Wei

The demand for high-quality medical imaging in clinical practice and assisted diagnosis has made 3D image reconstruction in radiological imaging a key research focus. Artificial intelligence (AI) has emerged as a promising approach for…

Computer Vision and Pattern Recognition · Computer Science 2026-04-29 Yuezhe Yang , Lei Bi , Boyu Yang , Yaqian Wang , Yang He , Yige Peng , Zhe Jin , Xingbo Dong , Jinman Kim

Compared to RGB semantic segmentation, RGBD semantic segmentation can achieve better performance by taking depth information into consideration. However, it is still problematic for contemporary segmenters to effectively exploit RGBD…

Computer Vision and Pattern Recognition · Computer Science 2019-05-27 Xinxin Hu , Kailun Yang , Lei Fei , Kaiwei Wang

Automated radiology report generation (RRG) aims to produce detailed textual reports from clinical imaging, such as computed tomography (CT) scans, to improve the accuracy and efficiency of diagnosis and provision of management advice. RRG…

Machine Learning · Computer Science 2025-07-03 Siyou Li , Pengyao Qin , Huanan Wu , Dong Nie , Arun J. Thirunavukarasu , Juntao Yu , Le Zhang

This work is an endeavor to develop a deep learning methodology for automated anatomical labeling of a given region of interest (ROI) in brain computed tomography (CT) scans. We combine both local and global context to obtain a…

Computer Vision and Pattern Recognition · Computer Science 2018-01-23 Srikrishna Varadarajan , Muktabh Mayank Srivastava , Monika Grewal , Pulkit Kumar

Face parsing infers a pixel-wise label to each facial component, which has drawn much attention recently. Previous methods have shown their success in face parsing, which however overlook the correlation among facial components. As a matter…

Computer Vision and Pattern Recognition · Computer Science 2021-10-13 Gusi Te , Wei Hu , Yinglu Liu , Hailin Shi , Tao Mei

Visual content and accompanied audio signals naturally formulate a joint representation to improve audio-visual (AV) related applications. While studies develop various AV representation learning frameworks, the importance of AV data…

Computer Vision and Pattern Recognition · Computer Science 2024-11-01 Shentong Mo , Yibing Song