English
Related papers

Related papers: Revisiting End-to-End Learning with Slide-level Su…

200 papers

Multiple Instance Learning (MIL) has emerged as the best solution for Whole Slide Image (WSI) classification. It consists of dividing each slide into patches, which are treated as a bag of instances labeled with a global label. MIL includes…

Computer Vision and Pattern Recognition · Computer Science 2025-05-05 Ali Mammadov , Loic Le Folgoc , Julien Adam , Anne Buronfosse , Gilles Hayem , Guillaume Hocquet , Pietro Gori

Convolutional neural networks can be trained to perform histology slide classification using weak annotations with multiple instance learning (MIL). However, given the paucity of labeled histology data, direct application of MIL can easily…

Computer Vision and Pattern Recognition · Computer Science 2019-11-05 Ming Y. Lu , Richard J. Chen , Jingwen Wang , Debora Dillon , Faisal Mahmood

Computational pathology holds substantial promise for improving diagnosis and guiding treatment decisions. Recent pathology foundation models enable the extraction of rich patch-level representations from large-scale whole-slide images…

Computer Vision and Pattern Recognition · Computer Science 2025-11-20 Xiangde Luo , Jinxi Xiang , Yuanfeng Ji , Ruijiang Li

End-to-end (E2E) autonomous driving models that take only camera images as input and directly predict a future trajectory are appealing for their computational efficiency and potential for improved generalization via unified optimization;…

Robotics · Computer Science 2026-04-10 Chihiro Noguchi , Takaki Yamamoto

Bag-based Multiple Instance Learning (MIL) approaches have emerged as the mainstream methodology for Whole Slide Image (WSI) classification. However, most existing methods adopt a segmented training strategy, which first extracts features…

Computer Vision and Pattern Recognition · Computer Science 2025-03-13 Jiangping Wen , Jinyu Wen , Meie Fang

In recent years, the availability of digitized Whole Slide Images (WSIs) has enabled the use of deep learning-based computer vision techniques for automated disease diagnosis. However, WSIs present unique computational and algorithmic…

Image and Video Processing · Electrical Eng. & Systems 2021-06-15 Yash Sharma , Aman Shrivastava , Lubaina Ehsan , Christopher A. Moskaluk , Sana Syed , Donald E. Brown

In autonomous driving, the end-to-end (E2E) driving approach that predicts vehicle control signals directly from sensor data is rapidly gaining attention. To learn a safe E2E driving system, one needs an extensive amount of driving data and…

Robotics · Computer Science 2025-05-12 Jin Bok Park , Jinkyu Lee , Muhyun Back , Hyunmin Han , David T. Ma , Sang Min Won , Sung Soo Hwang , Il Yong Chun

Whole-slide image (WSI) classification in computational pathology is commonly formulated as slide-level Multiple Instance Learning (MIL) with a single global bag representation. However, slide-level MIL is fundamentally underconstrained:…

Computer Vision and Pattern Recognition · Computer Science 2026-04-15 Syed Fahim Ahmed , Gnanesh Rasineni , Florian Koehler , Abu Zahid Bin Aziz , Mei Wang , Attila Gyulassy , Brian Summa , J. Quincy Brown , Valerio Pascucci , Shireen Y. Elhabian

Multiple Instance Learning (MIL) for whole slide image (WSI) analysis in computational pathology often neglects instance-level learning as supervision is typically provided only at the bag level, hindering the integrated consideration of…

Computer Vision and Pattern Recognition · Computer Science 2025-09-26 Shuyang Wu , Yifu Qiu , Ines P. Nearchou , Sandrine Prost , Jonathan A. Fallowfield , Hideki Ueno , Hitoshi Tsuda , David J. Harrison , Hakan Bilen , Timothy J. Kendall

Multiple Instance Learning is the predominant method for Whole Slide Image classification in digital pathology, enabling the use of slide-level labels to supervise model training. Although MIL eliminates the tedious fine-grained annotation…

Computer Vision and Pattern Recognition · Computer Science 2025-05-28 Chen Shu , Boyu Fu , Yiman Li , Ting Yin , Wenchuan Zhang , Jie Chen , Yuhao Yi , Hong Bu

End-to-End (E2E) unrolled optimization frameworks show promise for Magnetic Resonance (MR) image recovery, but suffer from high memory usage during training. In addition, these deterministic approaches do not offer opportunities for…

Image and Video Processing · Electrical Eng. & Systems 2024-02-09 Jyothi Rikhab Chand , Mathews Jacob

Graph-based Multiple Instance Learning (MIL) is widely used in survival analysis with Hematoxylin and Eosin (H\&E)-stained whole slide images (WSIs) due to its ability to capture topological information. However, variations in staining and…

Computer Vision and Pattern Recognition · Computer Science 2025-09-25 Min Cen , Zhenfeng Zhuang , Yuzhe Zhang , Min Zeng , Baptiste Magnier , Lequan Yu , Hong Zhang , Liansheng Wang

Masked image modeling (MIM) learns visual representation by masking and reconstructing image patches. Applying the reconstruction supervision on the CLIP representation has been proven effective for MIM. However, it is still under-explored…

Computer Vision and Pattern Recognition · Computer Science 2022-11-18 Xinyu Zhang , Jiahui Chen , Junkun Yuan , Qiang Chen , Jian Wang , Xiaodi Wang , Shumin Han , Xiaokang Chen , Jimin Pi , Kun Yao , Junyu Han , Errui Ding , Jingdong Wang

Self-supervised learning (SSL) has been successful in building patch embeddings of small histology images (e.g., 224x224 pixels), but scaling these models to learn slide embeddings from the entirety of giga-pixel whole-slide images (WSIs)…

Computer Vision and Pattern Recognition · Computer Science 2024-05-21 Guillaume Jaume , Lukas Oldenburg , Anurag Vaidya , Richard J. Chen , Drew F. K. Williamson , Thomas Peeters , Andrew H. Song , Faisal Mahmood

Artificial intelligence (AI) has transformed digital pathology by enabling biomarker prediction from high-resolution whole-slide images (WSIs). However, current methods are computationally inefficient, processing thousands of redundant…

This paper addresses the problem of end-to-end (E2E) design of learning and communication in a task-oriented semantic communication system. In particular, we consider a multi-device cooperative edge inference system over a wireless…

Information Theory · Computer Science 2024-09-02 Chang Cai , Xiaojun Yuan , Ying-Jun Angela Zhang

Masked Autoencoder (MAE) is a notable method for self-supervised pretraining in visual representation learning. It operates by randomly masking image patches and reconstructing these masked patches using the unmasked ones. A key limitation…

Computer Vision and Pattern Recognition · Computer Science 2025-03-25 Han Guo , Ramtin Hosseini , Ruiyi Zhang , Sai Ashish Somayajula , Ranak Roy Chowdhury , Rajesh K. Gupta , Pengtao Xie

Cooperative inference in Mobile Edge Computing (MEC), achieved by deploying partitioned Deep Neural Network (DNN) models between resource-constrained user equipments (UEs) and edge servers (ESs), has emerged as a promising paradigm.…

Distributed, Parallel, and Cluster Computing · Computer Science 2023-10-20 Xinrui Ye , Yanzan Sun , Dingzhu Wen , Guanjin Pan , Shunqing Zhang

Current approaches for classification of whole slide images (WSI) in digital pathology predominantly utilize a two-stage learning pipeline. The first stage identifies areas of interest (e.g. tumor tissue), while the second stage processes…

Computer Vision and Pattern Recognition · Computer Science 2022-07-20 Marvin Teichmann , Andre Aichert , Hanibal Bohnenberger , Philipp Ströbel , Tobias Heimann

Although end-to-end (E2E) learning has led to impressive progress on a variety of visual understanding tasks, it is often impeded by hardware constraints (e.g., GPU memory) and is prone to overfitting. When it comes to video captioning, one…

Computer Vision and Pattern Recognition · Computer Science 2019-01-03 Lijun Li , Boqing Gong