中文
相关论文

相关论文: Using Multi-modal Data for Improving Generalizabil…

200 篇论文

Understanding how novices acquire and hone visual search skills is crucial for developing and optimizing training methods across domains. Network analysis methods can be used to analyze graph representations of visual expertise. This study…

人机交互 · 计算机科学 2025-07-28 Pingjing Yang , Jennifer Cromley , Jana Diesner

Gaze estimation, which is a method to determine where a person is looking at given the person's full face, is a valuable clue for understanding human intention. Similarly to other domains of computer vision, deep learning (DL) methods have…

计算机视觉与模式识别 · 计算机科学 2022-08-09 Longzhao Huang , Yujie Li , Xu Wang , Haoyu Wang , Ahmed Bouridane , Ahmad Chaddad

Medical imaging is a very useful tool in healthcare, various technologies being employed to non-invasively peek inside the human body. Deep learning with neural networks in radiology was welcome - albeit cautiously - by the radiologist…

系统与控制 · 电气工程与系统科学 2024-06-18 Szilard Enyedi

Machine learning (ML) and deep learning (DL) techniques have been widely applied to analyze electroencephalography (EEG) signals for disease diagnosis and brain-computer interfaces (BCI). The integration of multimodal data has been shown to…

信号处理 · 电气工程与系统科学 2025-01-16 Siqi Zhao , Wangyang Li , Xiru Wang , Stevie Foglia , Hongzhao Tan , Bohan Zhang , Ameer Hamoodi , Aimee Nelson , Zhen Gao

Diabetic Retinopathy (DR) is a serious microvascular complication of diabetes, and one of the leading causes of vision loss worldwide. Although automated detection and grading, with Deep Learning (DL), can reduce the burden on…

图像与视频处理 · 电气工程与系统科学 2026-04-06 Shramana Dey , Zahir Khan , T. A. PramodKumar , B. Uma Shankar , Ashis K. Dhara , Ramachandran Rajalakshmi , Rajiv Raman , Sushmita Mitra

Eye tracking research is important in computer vision because it can help us understand how humans interact with the visual world. Specifically for high-risk applications, such as in medical imaging, eye tracking can help us to comprehend…

计算机视觉与模式识别 · 计算机科学 2023-08-31 Bin Wang , Hongyi Pan , Armstrong Aboah , Zheyuan Zhang , Elif Keles , Drew Torigian , Baris Turkbey , Elizabeth Krupinski , Jayaram Udupa , Ulas Bagci

Medical image segmentation aims to identify and locate abnormal structures in medical images, such as chest radiographs, using deep neural networks. These networks require a large number of annotated images with fine-grained masks for the…

图像与视频处理 · 电气工程与系统科学 2024-01-17 Jiamin Chen , Xuhong Li , Yanwu Xu , Mengnan Du , Haoyi Xiong

Advanced diagnostic instruments are crucial for the accurate detection and treatment of lung diseases, which affect millions of individuals globally. This study examines the effectiveness of deep learning and transfer learning models using…

图像与视频处理 · 电气工程与系统科学 2025-06-23 Shuvashis Sarker , Shamim Rahim Refat , Faika Fairuj Preotee , Tanvir Rouf Shawon , Raihan Tanvir

Medical image analysis often faces significant challenges due to limited expert-annotated data, hindering both model generalization and clinical adoption. We propose an expert-guided explainable few-shot learning framework that integrates…

图像与视频处理 · 电气工程与系统科学 2025-09-12 Ifrat Ikhtear Uddin , Longwei Wang , KC Santosh

Recent advancements in Computer Assisted Diagnosis have shown promising performance in medical imaging tasks, particularly in chest X-ray analysis. However, the interaction between these models and radiologists has been primarily limited to…

计算机视觉与模式识别 · 计算机科学 2024-04-04 Yunsoo Kim , Jinge Wu , Yusuf Abdulle , Yue Gao , Honghan Wu

In this paper, we consider the problem of disease diagnosis. Unlike the conventional learning paradigm that treats labels independently, we propose a knowledge-enhanced framework, that enables training visual representation with the…

计算机视觉与模式识别 · 计算机科学 2023-02-28 Chaoyi Wu , Xiaoman Zhang , Yanfeng Wang , Ya Zhang , Weidi Xie

Multimodal large models have shown great potential in automating pathology image analysis. However, current multimodal models for gastrointestinal pathology are constrained by both data quality and reasoning transparency: pervasive noise…

图像与视频处理 · 电气工程与系统科学 2025-07-25 Minxi Ouyang , Lianghui Zhu , Yaqing Bao , Qiang Huang , Jingli Ouyang , Tian Guan , Xitong Ling , Jiawen Li , Song Duan , Wenbin Dai , Li Zheng , Xuemei Zhang , Yonghong He

The manifold hypothesis is a core mechanism behind the success of deep learning, so understanding the intrinsic manifold structure of image data is central to studying how neural networks learn from the data. Intrinsic dataset manifolds and…

图像与视频处理 · 电气工程与系统科学 2022-09-19 Nicholas Konz , Hanxue Gu , Haoyu Dong , Maciej A. Mazurowski

Analyzing the gaze accuracy characteristics of an eye tracker is a critical task as its gaze data is frequently affected by non-ideal operating conditions in various consumer eye tracking applications. In this study, gaze error patterns…

信号处理 · 电气工程与系统科学 2020-05-11 Anuradha Kar

Vision-Language Models (VLMs) have demonstrated remarkable success in natural language generation, excelling at instruction following and structured output generation. Knowledge graphs play a crucial role in radiology, serving as valuable…

计算与语言 · 计算机科学 2025-05-26 Abdullah Abdullah , Seong Tae Kim

Deep Learning (DL) holds enormous potential for improving medical imaging diagnostics, yet the lack of interpretability in most models hampers clinical trust and adoption. This paper presents an explainable deep learning framework for…

计算机视觉与模式识别 · 计算机科学 2025-10-28 Sai Teja Erukude , Viswa Chaitanya Marella , Suhasnadh Reddy Veluru

Explainable Deep Learning has gained significant attention in the field of artificial intelligence (AI), particularly in domains such as medical imaging, where accurate and interpretable machine learning models are crucial for effective…

图像与视频处理 · 电气工程与系统科学 2024-09-11 Subhashis Suara , Aayush Jha , Pratik Sinha , Arif Ahmed Sekh

The medical field is creating large amount of data that physicians are unable to decipher and use efficiently. Moreover, rule-based expert systems are inefficient in solving complicated medical tasks or for creating insights using big data.…

计算机视觉与模式识别 · 计算机科学 2024-04-05 Paschalis Bizopoulos , Dimitrios Koutsouris

Automatic radiology report generation is a promising application of multimodal deep learning, aiming to reduce reporting workload and improve consistency. However, current state-of-the-art (SOTA) systems - such as Multimodal AI for…

Curating a large scale medical imaging dataset for machine learning applications is both time consuming and expensive. Balancing the workload between model development, data collection and annotations is difficult for machine learning…

人工智能 · 计算机科学 2022-06-07 Athanasios Vlontzos , Hadrien Reynaud , Bernhard Kainz