English
Related papers

Related papers: Transformer-based Model for Oral Epithelial Dyspla…

200 papers

Synthetic PET images are valuable for quantitative imaging workflow development, scalable virtual imaging trials, and deep learning model training, but conventional physics-based simulation approaches are computationally intensive, limited…

Computer Vision and Pattern Recognition · Computer Science 2026-05-21 Suya Li , Kaushik Dutta , Debojyoti Pal , Jingqin Luo , Kooresh I. Shoghi

Deep generative models have been demonstrated as problematic in the unsupervised out-of-distribution (OOD) detection task, where they tend to assign higher likelihoods to OOD samples. Previous studies on this issue are usually not…

Machine Learning · Computer Science 2024-01-04 Zezhen Zeng , Bin Liu

Transfer learning has gained attention in medical image analysis due to limited annotated 3D medical datasets for training data-driven deep learning models in the real world. Existing 3D-based methods have transferred the pre-trained models…

Computer Vision and Pattern Recognition · Computer Science 2021-04-29 Eunji Jun , Seungwoo Jeong , Da-Woon Heo , Heung-Il Suk

Eosinophilic esophagitis (EoE) is a chronic allergic inflammatory condition of the esophagus associated with elevated esophageal eosinophils. Second only to gastroesophageal reflux disease, EoE is one of the leading causes of chronic…

We develop a transformer-based sequence-to-sequence model that recovers scalar ordinary differential equations (ODEs) in symbolic form from irregularly sampled and noisy observations of a single solution trajectory. We demonstrate in…

Machine Learning · Computer Science 2023-07-25 Sören Becker , Michal Klein , Alexander Neitz , Giambattista Parascandolo , Niki Kilbertus

We present OffRoadTranSeg, the first end-to-end framework for semi-supervised segmentation in unstructured outdoor environment using transformers and automatic data selection for labelling. The offroad segmentation is a scene understanding…

Computer Vision and Pattern Recognition · Computer Science 2021-06-29 Anukriti Singh , Kartikeya Singh , P. B. Sujit

Transformer architectures, including nnFormer,have demonstrated promising results in volumetric medical image segmentation by being able to capture long-range spatial interactions. Although they have high performance, these models need…

Computer Vision and Pattern Recognition · Computer Science 2026-04-28 R. M. Krishna Sureddi , T. Satyanarayana Murthy , Nomula Varsha Reddy , Adi Kanishka , Nalla Manvika Reddy

Transformer-based deep learning models have demonstrated exceptional performance in medical imaging by leveraging attention mechanisms for feature representation and interpretability. However, these models are prone to learning spurious…

Computer Vision and Pattern Recognition · Computer Science 2025-10-15 Shelley Zixin Shu , Haozhe Luo , Alexander Poellinger , Mauricio Reyes

Purpose: To present a high-performing, robust, and flexible deep learning pipeline for automatic segmentation of 30 organs-at-risk (OARs) in head and neck (H&N) cancer patients, using MRI, CT, or both. Method: We trained a segmentation…

Image and Video Processing · Electrical Eng. & Systems 2025-09-08 Sébastien Quetin , Andrew Heschl , Mauricio Murillo , Rohit Murali , Piotr Pater , George Shenouda , Shirin A. Enger , Farhad Maleki

Models trained with deep learning often fail to signal when inputs fall outside their training data manifold, leading to unreliable predictions under distribution shift. Prior work suggests that effective out-of-distribution (OOD) detection…

Machine Learning · Computer Science 2026-05-08 Brett Barkley , Preston Culbertson , David Fridovich-Keil

Periorbital distances are critical markers for diagnosing and monitoring a range of oculoplastic and craniofacial conditions. Manual measurement, however, is subjective and prone to intergrader variability. Automated methods have been…

Motivation: Deep learning models deployed for use on medical tasks can be equipped with Out-of-Distribution Detection (OoDD) methods in order to avoid erroneous predictions. However it is unclear which OoDD method should be used in…

Machine Learning · Computer Science 2020-08-06 Tianshi Cao , Chin-Wei Huang , David Yu-Tung Hui , Joseph Paul Cohen

Vision Transformers are at the heart of the current surge of interest in foundation models for histopathology. They process images by breaking them into smaller patches following a regular grid, regardless of their content. Yet, not all…

Computer Vision and Pattern Recognition · Computer Science 2024-04-30 Clément Grisi , Geert Litjens , Jeroen van der Laak

3D medical image segmentation is a challenging task with crucial implications for disease diagnosis and treatment planning. Recent advances in deep learning have significantly enhanced fully supervised medical image segmentation. However,…

Image and Video Processing · Electrical Eng. & Systems 2025-06-23 Runmin Jiang , Zhaoxin Fan , Junhao Wu , Lenghan Zhu , Xin Huang , Tianyang Wang , Heng Huang , Min Xu

Recently Transformer and Convolution neural network (CNN) based models have shown promising results in EEG signal processing. Transformer models can capture the global dependencies in EEG signals through a self-attention mechanism, while…

Signal Processing · Electrical Eng. & Systems 2023-09-11 Chenyu Liu , Xinliang Zhou , Yang Liu

Optic nerve head (ONH) detection has been a crucial area of study in ophthalmology for years. However, the significant discrepancy between fundus image datasets, each generated using a single type of fundus camera, poses challenges to the…

Image and Video Processing · Electrical Eng. & Systems 2024-06-04 Jiayi Wang , Yi-An Mao , Xiaoyu Ma , Sicen Guo , Yuting Shao , Xiao Lv , Wenting Han , Mark Christopher , Linda M. Zangwill , Yanlong Bi , Rui Fan

Purpose: Trochlear Dysplasia (TD) is a common malformation in adolescents, leading to anterior knee pain and instability. Surgical interventions such as trochleoplasty require precise planning to correct the trochlear groove. However, no…

Image and Video Processing · Electrical Eng. & Systems 2024-12-16 Michael Wehrli , Alicia Durrer , Paul Friedrich , Volodimir Buchakchiyskiy , Marcus Mumme , Edwin Li , Gyozo Lehoczky , Carol C. Hasler , Philippe C. Cattin

Multi-modal foundation models are typically trained on millions of pairs of natural images and text captions, frequently obtained through web-crawling approaches. Although such models depict excellent generative capabilities, they do not…

Computer Vision and Pattern Recognition · Computer Science 2023-01-03 Pierre Chambon , Christian Bluethgen , Curtis P. Langlotz , Akshay Chaudhari

Positron emission tomography (PET) is an advanced medical imaging technique that plays a crucial role in non-invasive clinical diagnosis. However, while reducing radiation exposure through low-dose PET scans is beneficial for patient…

Computer Vision and Pattern Recognition · Computer Science 2024-07-02 Bin Huang , Xubiao Liu , Lei Fang , Qiegen Liu , Bingxuan Li

Deep learning-based image segmentation and detection models have largely improved the efficiency of analyzing retinal landmarks such as optic disc (OD), optic cup (OC), and fovea. However, factors including ophthalmic disease-related…

Image and Video Processing · Electrical Eng. & Systems 2023-05-22 Huaqing He , Li Lin , Zhiyuan Cai , Pujin Cheng , Xiaoying Tang