中文
相关论文

相关论文: Cross-modal Prototype Driven Network for Radiology…

200 篇论文

In clinical scenarios, multi-specialist consultation could significantly benefit the diagnosis, especially for intricate cases. This inspires us to explore a "multi-expert joint diagnosis" mechanism to upgrade the existing "single expert"…

计算机视觉与模式识别 · 计算机科学 2023-04-06 Zhanyu Wang , Lingqiao Liu , Lei Wang , Luping Zhou

We propose MARL-Rad, a multi-modal multi-agent reinforcement learning framework for radiology report generation that trains the entire agentic system on policy within its deployed radiology workflow. MARL-Rad addresses the limitation of…

计算机视觉与模式识别 · 计算机科学 2026-05-11 Kaito Baba , Risa Kishikawa , Satoshi Kodera

Survival prediction is a crucial task in the medical field and is essential for optimizing treatment options and resource allocation. However, current methods often rely on limited data modalities, resulting in suboptimal performance. In…

图像与视频处理 · 电气工程与系统科学 2025-01-07 Binyu Zhang , Zhu Meng , Junhao Dong , Fei Su , Zhicheng Zhao

Generating radiology reports is time-consuming and requires extensive expertise in practice. Therefore, reliable automatic radiology report generation is highly desired to alleviate the workload. Although deep learning techniques have been…

图像与视频处理 · 电气工程与系统科学 2019-07-24 Jianbo Yuan , Haofu Liao , Rui Luo , Jiebo Luo

Automated radiology report generation from chest X-ray (CXR) images has the potential to improve clinical efficiency and reduce radiologists' workload. However, most datasets, including the publicly available MIMIC-CXR and CheXpert Plus,…

Automatic radiology report generation has attracted enormous research interest due to its practical value in reducing the workload of radiologists. However, simultaneously establishing global correspondences between the image (e.g., Chest…

计算机视觉与模式识别 · 计算机科学 2023-07-19 Yaowei Li , Bang Yang , Xuxin Cheng , Zhihong Zhu , Hongxiang Li , Yuexian Zou

Automatic generation of ophthalmic reports using data-driven neural networks has great potential in clinical practice. When writing a report, ophthalmologists make inferences with prior clinical knowledge. This knowledge has been neglected…

计算机视觉与模式识别 · 计算机科学 2022-06-07 Mingjie Li , Wenjia Cai , Karin Verspoor , Shirui Pan , Xiaodan Liang , Xiaojun Chang

Multimodal molecular representation learning, which jointly models molecular graphs and their textual descriptions, enhances predictive accuracy and interpretability by enabling more robust and reliable predictions of drug toxicity,…

机器学习 · 计算机科学 2025-10-21 Yingxu Wang , Kunyu Zhang , Jiaxin Huang , Nan Yin , Siwei Liu , Eran Segal

Recent advancements in artificial intelligence have significantly improved the automatic generation of radiology reports. However, existing evaluation methods fail to reveal the models' understanding of radiological images and their…

人工智能 · 计算机科学 2024-08-27 Xiaoman Zhang , Julián N. Acosta , Hong-Yu Zhou , Pranav Rajpurkar

Labeling training datasets has become a key barrier to building medical machine learning models. One strategy is to generate training labels programmatically, for example by applying natural language processing pipelines to text reports…

Cancer survival prediction requires integrating pathological Whole Slide Images (WSIs) and genomic profiles, a challenging task due to the inherent heterogeneity and the complexity of modeling both inter- and intra-modality interactions.…

图像与视频处理 · 电气工程与系统科学 2025-07-08 Mingxin Liu , Chengfei Cai , Jun Li , Pengbo Xu , Jinze Li , Jiquan Ma , Jun Xu

The extraction of structured clinical information from free-text radiology reports in the form of radiology graphs has been demonstrated to be a valuable approach for evaluating the clinical correctness of report-generation methods.…

计算机视觉与模式识别 · 计算机科学 2023-09-19 Yiheng Xiong , Jingsong Liu , Kamilia Zaripova , Sahand Sharifzadeh , Matthias Keicher , Nassir Navab

With the aim of matching a pair of instances from two different modalities, cross modality mapping has attracted growing attention in the computer vision community. Existing methods usually formulate the mapping function as the similarity…

计算机视觉与模式识别 · 计算机科学 2021-03-01 Zun Li , Congyan Lang , Liqian Liang , Tao Wang , Songhe Feng , Jun Wu , Yidong Li

Radiology report generation (RRG) for diagnostic images, such as chest X-rays, plays a pivotal role in both clinical practice and AI. Traditional free-text reports suffer from redundancy and inconsistent language, complicating the…

计算机视觉与模式识别 · 计算机科学 2025-08-05 Yingshu Li , Yunyi Liu , Zhanyu Wang , Xinyu Liang , Lingqiao Liu , Lei Wang , Luping Zhou

Histo-genomic multimodal survival prediction has garnered growing attention for its remarkable model performance and potential contributions to precision medicine. However, a significant challenge in clinical practice arises when only…

机器学习 · 计算机科学 2025-03-17 Fengchun Liu , Linghan Cai , Zhikang Wang , Zhiyuan Fan , Jin-gang Yu , Hao Chen , Yongbing Zhang

Structured radiology reporting promises faster, more consistent communication than free text, but automation remains difficult as models must make many fine-grained, discrete decisions about rare findings and attributes from limited…

人工智能 · 计算机科学 2026-03-13 Chantal Pellegrini , Adrian Delchev , Ege Özsoy , Nassir Navab , Matthias Keicher

To reduce doctors' workload, deep-learning-based automatic medical report generation has recently attracted more and more research efforts, where attention mechanisms and reinforcement learning are integrated with the classic…

计算与语言 · 计算机科学 2020-11-17 Wenting Xu , Chang Qi , Zhenghua Xu , Thomas Lukasiewicz

Drafting radiology reports is a complex task requiring flexibility, where radiologists tail content to available information and particular clinical demands. However, most current radiology report generation (RRG) models are constrained to…

计算与语言 · 计算机科学 2024-12-17 Zhuhao Wang , Yihua Sun , Zihan Li , Xuan Yang , Fang Chen , Hongen Liao

Current methods for few-shot action recognition mainly fall into the metric learning framework following ProtoNet, which demonstrates the importance of prototypes. Although they achieve relatively good performance, the effect of multimodal…

计算机视觉与模式识别 · 计算机科学 2024-05-22 Xinzhe Ni , Yong Liu , Hao Wen , Yatai Ji , Jing Xiao , Yujiu Yang

Electroencephalography(EEG)-basedemotionrecognitionre- mains challenging in cross-subject settings due to severe inter-subject variability. Existing methods mainly learn subject-invariant features, but often under-exploit stimulus-locked…

机器学习 · 计算机科学 2026-03-13 Renwei Meng