中文
相关论文

相关论文: I-ODA, Real-World Multi-modal Longitudinal Data fo…

200 篇论文

Safety and efficiency are paramount in healthcare facilities where the lives of patients are at stake. Despite the adoption of robots to assist medical staff in challenging tasks such as complex surgeries, human expertise is still…

计算机视觉与模式识别 · 计算机科学 2024-12-11 Rohit Mohan , José Arce , Sassan Mokhtar , Daniele Cattaneo , Abhinav Valada

Artificial intelligence has demonstrated significant potential in clinical decision-making; however, developing models capable of adapting to diverse real-world scenarios and performing complex diagnostic reasoning remains a major…

计算机视觉与模式识别 · 计算机科学 2025-08-18 Ronghao Xu , Zhen Huang , Yangbo Wei , Xiaoqian Zhou , Zikang Xu , Ting Liu , Zihang Jiang , S. Kevin Zhou

While deep learning methods have shown great success in medical image analysis, they require a number of medical images to train. Due to data privacy concerns and unavailability of medical annotators, it is oftentimes very difficult to…

图像与视频处理 · 电气工程与系统科学 2020-10-08 Yue Yang , Pengtao Xie

This paper introduces MedTrinity-25M, a comprehensive, large-scale multimodal dataset for medicine, covering over 25 million images across 10 modalities with multigranular annotations for more than 65 diseases. These multigranular…

计算机视觉与模式识别 · 计算机科学 2025-07-11 Yunfei Xie , Ce Zhou , Lang Gao , Juncheng Wu , Xianhang Li , Hong-Yu Zhou , Sheng Liu , Lei Xing , James Zou , Cihang Xie , Yuyin Zhou

Recent technological advances in healthcare have led to unprecedented growth in patient data quantity and diversity. While artificial intelligence (AI) models have shown promising results in analyzing individual data modalities, there is…

Deep learning has recently gained high interest in ophthalmology, due to its ability to detect clinically significant features for diagnosis and prognosis. Despite these significant advances, little is known about the ability of various…

计算机视觉与模式识别 · 计算机科学 2019-05-28 Petteri Teikari , Raymond P. Najjar , Leopold Schmetterer , Dan Milea

While the field of medical image analysis has undergone a transformative shift with the integration of machine learning techniques, the main challenge of these techniques is often the scarcity of large, diverse, and well-annotated datasets.…

计算机视觉与模式识别 · 计算机科学 2025-11-18 Stefano Woerner , Arthur Jaques , Christian F. Baumgartner

We present a large scale data set, OpenEDS: Open Eye Dataset, of eye-images captured using a virtual-reality (VR) head mounted display mounted with two synchronized eyefacing cameras at a frame rate of 200 Hz under controlled illumination.…

计算机视觉与模式识别 · 计算机科学 2019-05-20 Stephan J. Garbin , Yiru Shen , Immo Schuetz , Robert Cavin , Gregory Hughes , Sachin S. Talathi

Diversity-aware data are essential for a robust modeling of human behavior in context. In addition, being the human behavior of interest for numerous applications, data must also be reusable across domain, to ensure diversity of…

计算机与社会 · 计算机科学 2023-06-19 Matteo Busso , Xiaoyue Li

As integrated circuit (IC) dimensions shrink below the lithographic wavelength, optical lithography faces growing challenges from diffraction and process variability. Model-based optical proximity correction (OPC) and inverse lithography…

机器学习 · 计算机科学 2025-12-25 Yuting Hu , Lei Zhuang , Hua Xiang , Jinjun Xiong , Gi-Joon Nam

This study presents a dataset consisting of 268 retinal images from 179 individuals, including 133 left-eye and 135 right-eye images, collected from Natasha Eye Care and Research Institute in Pune, Maharashtra, India. The images were…

图像与视频处理 · 电气工程与系统科学 2024-09-09 Pooja Bidwai , Shilpa Gite , Biswajeet Pradhan , Aditi Gupta , Kishore pahuja

Industrial anomaly detection (IAD) has garnered significant attention and experienced rapid development. However, the recent development of IAD approach has encountered certain difficulties due to dataset limitations. On the one hand, most…

计算机视觉与模式识别 · 计算机科学 2024-03-20 Chengjie Wang , Wenbing Zhu , Bin-Bin Gao , Zhenye Gan , Jianning Zhang , Zhihao Gu , Shuguang Qian , Mingang Chen , Lizhuang Ma

As HPC systems grow in complexity, efficient and manageable operation is increasingly critical. Many centers are thus starting to explore the use of Operational Data Analytics (ODA) techniques, which extract knowledge from massive amounts…

分布式、并行与集群计算 · 计算机科学 2021-06-29 Alessio Netti , Michael Ott , Carla Guillen , Daniele Tafani , Martin Schulz

Large-scale multimodal models achieve strong results on tasks like Visual Question Answering (VQA), but they are often limited when queries require cultural and visual information, everyday knowledge, particularly in low-resource and…

Despite significant advances in artificial intelligence (AI) for computer vision, its application in medical imaging has been limited by the burden and limits of expert-generated labels. We used images from optical coherence tomography…

计算机视觉与模式识别 · 计算机科学 2018-02-27 Cecilia S. Lee , Ariel J. Tyring , Yue Wu , Sa Xiao , Ariel S. Rokem , Nicolaas P. Deruyter , Qinqin Zhang , Adnan Tufail , Ruikang K. Wang , Aaron Y. Lee

Recent advancements in Digital Pathology (DP), particularly through artificial intelligence and Foundation Models, have underscored the importance of large-scale, diverse, and richly annotated datasets. Despite their critical role, publicly…

图像与视频处理 · 电气工程与系统科学 2025-05-20 Dmitry Nechaev , Alexey Pchelnikov , Ekaterina Ivanova

We propose a method called integrated diffusion for combining multimodal datasets, or data gathered via several different measurements on the same system, to create a joint data diffusion operator. As real world data suffers from both local…

机器学习 · 计算机科学 2022-03-07 Manik Kuchroo , Abhinav Godavarthi , Alexander Tong , Guy Wolf , Smita Krishnaswamy

We present a new dataset with annotated eye movements. The dataset consists of over 800,000 gaze points recorded during a car ride in the real world and in the simulator. In total, the eye movements of 19 subjects were annotated. In this…

计算机视觉与模式识别 · 计算机科学 2021-01-13 Wolfgang Fuhl , Enkelejda Kasneci

We introduce OLATverse, a large-scale dataset comprising around 9M images of 765 real-world objects, captured from multiple viewpoints under a diverse set of precisely controlled lighting conditions. While recent advances in object-centric…

Recently, Vision-Language Models (VLMs) have achieved remarkable progress in multimodal tasks, and multimodal instruction data serves as the foundation for enhancing VLM capabilities. Despite the availability of several open-source…