中文
相关论文

相关论文: HiPath: Hierarchical Vision-Language Alignment for…

200 篇论文

Medical large vision-language Models (Med-LVLMs) have shown promise in clinical applications but suffer from factual inaccuracies and unreliable outputs, posing risks in real-world diagnostics. While RAG has emerged as a potential solution,…

计算与语言 · 计算机科学 2026-05-05 Zhe Chen , Yusheng Liao , Zhiyuan Zhu , Haolin Li , Hongcheng Liu , Yanfeng Wang , Yu Wang

Computational pathology (CoPath) leverages histopathology images to enhance diagnostic precision and reproducibility in clinical pathology. However, publicly available datasets for CoPath that are annotated with extensive histological…

The rapid digitization of histopathology slides has opened up new possibilities for computational tools in clinical and research workflows. Among these, content-based slide retrieval stands out, enabling pathologists to identify…

计算机视觉与模式识别 · 计算机科学 2025-10-28 Hongyi Wang , Zhengjie Zhu , Jiabo Ma , Fang Wang , Yue Shi , Bo Luo , Jili Wang , Qiuyu Cai , Xiuming Zhang , Yen-Wei Chen , Lanfen Lin , Hao Chen

While existing hierarchical text classification (HTC) methods attempt to capture label hierarchies for model training, they either make local decisions regarding each label or completely ignore the hierarchy information during inference. To…

信息检索 · 计算机科学 2020-06-19 Yuning Mao , Jingjing Tian , Jiawei Han , Xiang Ren

We present our submission to Task 3 (Discourse Relation Classification) of the DISRPT 2025 shared task. Task 3 introduces a unified set of 17 discourse relation labels across 39 corpora in 16 languages and six discourse frameworks, posing…

计算与语言 · 计算机科学 2025-09-23 Nawar Turk , Daniele Comitogianni , Leila Kosseim

This paper presents a novel holistic deep learning framework that simultaneously addresses the challenges of vulnerability to input perturbations, overparametrization, and performance instability from different train-validation splits. The…

The field of computational pathology has witnessed remarkable progress in the development of both task-specific predictive models and task-agnostic self-supervised vision encoders. However, despite the explosive growth of generative…

计算机视觉与模式识别 · 计算机科学 2023-12-14 Ming Y. Lu , Bowen Chen , Drew F. K. Williamson , Richard J. Chen , Kenji Ikamura , Georg Gerber , Ivy Liang , Long Phi Le , Tong Ding , Anil V Parwani , Faisal Mahmood

Vision Language Models (VLMs) have undergone significant advancements, particularly with the emergence of mobile-oriented VLMs, which offer a wide range of application scenarios. However, the substantial computational requirements for…

机器学习 · 计算机科学 2025-12-25 Yuanhao Xi , Xiaohuan Bing , Ramin Yahyapour

Vision language models (VLM) have achieved success in both natural language comprehension and image recognition tasks. However, their use in pathology report generation for whole slide images (WSIs) is still limited due to the huge size of…

计算机视觉与模式识别 · 计算机科学 2024-09-25 Jing Wei Tan , SeungKyu Kim , Eunsu Kim , Sung Hak Lee , Sangjeong Ahn , Won-Ki Jeong

Histopathology, the microscopic study of diseased tissue, is increasingly digitized, enabling improved visualization and streamlined workflows. An important task in histopathology is the segmentation of cells and glands, essential for…

计算机视觉与模式识别 · 计算机科学 2024-12-02 Philipp Endres , Valentin Koch , Julia A. Schnabel , Carsten Marr

Recent rapid progress in the field of computational pathology has been enabled by foundation models. These models are beginning to move beyond encoding image patches towards whole-slide understanding but their clinical utility remains…

Identifying cell types and subtypes in routine histopathology is fundamental for understanding disease. Existing tile-based models capture nuclear detail but miss the broader tissue context that influences cell identity. Current human…

计算机视觉与模式识别 · 计算机科学 2025-11-20 Yinuo Xu , Yan Cui , Mingyao Li , Zhi Huang

Vision-Language Models (VLMs) encode images and videos into abundant tokens, which contain substantial redundancy and computation cost. While visual token pruning mitigates the issue, most existing methods lack insight into the intrinsic…

计算机视觉与模式识别 · 计算机科学 2026-04-21 Jizhihui Liu , Feiyi Du , Guangdao Zhu , Niu Lian , Jun Li , Bin Chen , Weili Guan , Yaowei Wang

Large Language Models (LLMs) are increasingly deployed across diverse domains, raising the need for rigorous reliability assessment methods. Existing benchmark-based evaluations primarily offer descriptive statistics of model accuracy over…

软件工程 · 计算机科学 2026-01-30 Robab Aghazadeh-Chakherlou , Qing Guo , Siddartha Khastgir , Peter Popov , Xiaoge Zhang , Xingyu Zhao

Coarse-to-fine path decision-making requires predicting a valid taxonomy path in which earlier decisions constrain later ones. However, existing benchmarks score each level independently, obscuring cross-level validity and consistency. To…

多媒体 · 计算机科学 2026-03-20 Wei Yang , Yiran Zhu , Zilin Li , Xunjia Zhang , Jun Xia , Hongtao Wang

In pathology, the spatial distribution and proportions of tissue types are key indicators of disease progression, and are more readily available than fine-grained annotations. However, these assessments are rarely mapped to pixel-wise…

图像与视频处理 · 电气工程与系统科学 2026-04-28 Yangping Li , Thomas Pinetz , Michael Hölzel , Marieta Toma , Alexander Effland

The limited capacity for fine-grained visual perception presents a critical bottleneck for Vision-Language Models (VLMs) in real-world applications. Addressing this is challenging due to the scarcity of high-quality data and the limitations…

计算机视觉与模式识别 · 计算机科学 2025-10-29 Juntian Zhang , Song Jin , Chuanqi Cheng , Yuhan Liu , Yankai Lin , Xun Zhang , Yufei Zhang , Fei Jiang , Guojun Yin , Wei Lin , Rui Yan

Radiology reporting is a crucial part of the communication between radiologists and other medical professionals, but it can be time-consuming and error-prone. One approach to alleviate this is structured reporting, which saves time and…

计算机视觉与模式识别 · 计算机科学 2023-09-08 Chantal Pellegrini , Matthias Keicher , Ege Özsoy , Nassir Navab

Large Language Models (LLMs) are increasingly proposed for clinical decision support including multilingual diagnosis in low-resource settings. However, their reliability, calibration and safety characteristics remain insufficiently…

计算与语言 · 计算机科学 2026-05-05 Danish Ali , Li Xiaojian , Sundas Iqbal , Farrukh Zaidi

Pretraining has proven to be a powerful technique in natural language processing (NLP), exhibiting remarkable success in various NLP downstream tasks. However, in the medical domain, existing pretrained models on electronic health records…

人工智能 · 计算机科学 2023-10-23 Xiaochen Wang , Junyu Luo , Jiaqi Wang , Ziyi Yin , Suhan Cui , Yuan Zhong , Yaqing Wang , Fenglong Ma