English
Related papers

Related papers: Revisiting Automatic Data Curation for Vision Foun…

200 papers

Foundation Models (FMs) have shown impressive performance on various text and image processing tasks. They can generalize across domains and datasets in a zero-shot setting. This could make them suitable for automated quality inspection…

Computer Vision and Pattern Recognition · Computer Science 2025-09-26 Simon Baeuerle , Pratik Khanna , Nils Friederich , Angelo Jovin Yamachui Sitcheu , Damir Shakirov , Andreas Steimer , Ralf Mikut

Harnessing the intrinsic dynamics of physical systems for information processing opens new avenues for computation embodied in matter. Using simulations of a model system, we show that assemblies of DNA tiles capable of self-organizing into…

Soft Condensed Matter · Physics 2025-10-23 Tim E. Veenstra , René van Roij , Marjolein Dijkstra

Whole slide imaging (WSI) has transformed digital pathology by enabling computational analysis of gigapixel histopathology images. Recent foundation model advances have accelerated progress in computational pathology, facilitating joint…

Computer Vision and Pattern Recognition · Computer Science 2026-03-20 Peihang Wu , Zehong Chen , Lijian Xu

The deployment of foundation models for medical imaging has demonstrated considerable success. However, their training overheads associated with downstream tasks remain substantial due to the size of the image encoders employed, and the…

Computer Vision and Pattern Recognition · Computer Science 2025-04-04 Chengxi Zeng , Yuxuan Jiang , Fan Zhang , Alberto Gambaruto , Tilo Burghardt

Self-supervised pre-training has become the priory choice to establish reliable neural networks for automated recognition of massive biomedical microscopy images, which are routinely annotation-free, without semantics, and without guarantee…

Computer Vision and Pattern Recognition · Computer Science 2023-01-13 Wei Chen , Chen Li , Dan Chen , Xin Luo

Web-scraped, in-the-wild datasets have become the norm in face recognition research. The numbers of subjects and images acquired in web-scraped datasets are usually very large, with number of images on the millions scale. A variety of…

Computer Vision and Pattern Recognition · Computer Science 2020-04-08 Kai Zhang , Vítor Albiero , Kevin W. Bowyer

We evaluate the performance of federated learning (FL) in developing deep learning models for analysis of digitized tissue sections. A classification application was considered as the example use case, on quantifiying the distribution of…

Image and Video Processing · Electrical Eng. & Systems 2022-04-04 Ujjwal Baid , Sarthak Pati , Tahsin M. Kurc , Rajarsi Gupta , Erich Bremer , Shahira Abousamra , Siddhesh P. Thakur , Joel H. Saltz , Spyridon Bakas

The advent of foundation models (FMs) such as large language models (LLMs) has led to a cultural shift in data science, both in medicine and beyond. This shift involves moving away from specialized predictive models trained for specific,…

Machine Learning · Computer Science 2024-09-18 Ahmed Alaa , Bin Yu

Foundation models, first introduced in 2021, refer to large-scale pretrained models (e.g., large language models (LLMs) and vision-language models (VLMs)) that learn from extensive unlabeled datasets through unsupervised methods, enabling…

Deep learning methods are widely used for medical applications to assist medical doctors in their daily routines. While performances reach expert's level, interpretability (highlight how and what a trained model learned and why it makes a…

Computer Vision and Pattern Recognition · Computer Science 2020-09-30 Antoine Pirovano , Hippolyte Heuberger , Sylvain Berlemont , Saïd Ladjal , Isabelle Bloch

Machine learning algorithms underpin modern diagnostic-aiding software, which has proved valuable in clinical practice, particularly in radiology. However, inaccuracies, mainly due to the limited availability of clinical samples for…

Medical vision-language models (Med-VLMs) have shown impressive results in tasks such as report generation and visual question answering, but they still face several limitations. Most notably, they underutilize patient metadata and lack…

Computer Vision and Pattern Recognition · Computer Science 2025-10-16 Fangqi Cheng , Surajit Ray , Xiaochen Yang

AI-based biomarkers can infer molecular features directly from hematoxylin & eosin (H&E) slides, yet most pathology foundation models (PFMs) rely on global patch-level embeddings and overlook cell-level morphology. We present a PFM model,…

Computer Vision and Pattern Recognition · Computer Science 2025-11-10 Jingsong Liu , Han Li , Nassir Navab , Peter J. Schüffler

Osteoporosis is a common condition that increases fracture risk, especially in older adults. Early diagnosis is vital for preventing fractures, reducing treatment costs, and preserving mobility. However, healthcare providers face challenges…

Computer Vision and Pattern Recognition · Computer Science 2025-10-21 Mehdi Hosseini Chagahi , Saeed Mohammadi Dashtaki , Niloufar Delfan , Nadia Mohammadi , Farshid Rostami Pouria , Behzad Moshiri , Md. Jalil Piran , Oliver Faust

Whole Slide Imaging (WSI) has become a gold standard in cancer diagnosis, inspecting multi-scale information from cellular to tissue levels. Processing an entire WSI directly is infeasible due to GPU memory constraints; thus, Multiple…

Image and Video Processing · Electrical Eng. & Systems 2026-05-08 Tianyi Zhang , Sicheng Chen , Borui Kang , Dankai Liao , Qiaochu Xue , Bochong Zhang , Fei Xia , Zeyu Liu , Yueming Jin

Foundation models (FMs) promise to generalize medical imaging, but their effectiveness varies. It remains unclear how pre-training domain (medical vs. general), paradigm (e.g., text-guided), and architecture influence embedding quality,…

Fourier ptychographic microscopy (FPM) is a computational approach geared towards creating high-resolution and large field-of-view images without mechanical scanning. To acquire color images of histology slides, it often requires sequential…

Image and Video Processing · Electrical Eng. & Systems 2020-10-28 Ruihai Wang , Pengming Song , Shaowei Jiang , Chenggang Yan , Jiakai Zhu , Chengfei Guo , Zichao Bian , Tianbo Wang , Guoan Zheng

Nucleus segmentation is an important analysis task in digital pathology. However, methods for automatic segmentation often struggle with new data from a different distribution, requiring users to manually annotate nuclei and retrain…

Image and Video Processing · Electrical Eng. & Systems 2025-06-03 Titus Griebel , Anwai Archit , Constantin Pape

Pixel-aligned implicit models, such as PIFu, PIFuHD, and ICON, are used for single-view clothed human reconstruction. These models need to be trained using a sampling training scheme. Existing sampling training schemes either fail to…

Computer Vision and Pattern Recognition · Computer Science 2024-11-12 Kennard Yanting Chan , Fayao Liu , Guosheng Lin , Chuan Sheng Foo , Weisi Lin

Vision-language models (VLMs) are trained for thousands of GPU hours on carefully curated web datasets. In recent times, data curation has gained prominence with several works developing strategies to retain 'high-quality' subsets of 'raw'…

Machine Learning · Computer Science 2024-04-11 Sachin Goyal , Pratyush Maini , Zachary C. Lipton , Aditi Raghunathan , J. Zico Kolter
‹ Prev 1 8 9 10 Next ›