English
Related papers

Related papers: OPTED: Open Preprocessed Trachoma Eye Dataset Usin…

200 papers

3D part segmentation is a crucial and challenging task in 3D perception, playing a vital role in applications such as robotics, 3D generation, and 3D editing. Recent methods harness the powerful Vision Language Models (VLMs) for 2D-to-3D…

Computer Vision and Pattern Recognition · Computer Science 2024-11-19 Yunhan Yang , Yukun Huang , Yuan-Chen Guo , Liangjun Lu , Xiaoyang Wu , Edmund Y. Lam , Yan-Pei Cao , Xihui Liu

Medical image segmentation of anatomical structures and pathology is crucial in modern clinical diagnosis, disease study, and treatment planning. To date, great progress has been made in deep learning-based segmentation techniques, but most…

Computer Vision and Pattern Recognition · Computer Science 2024-06-21 Taha Koleilat , Hojat Asgariandehkordi , Hassan Rivaz , Yiming Xiao

We present SAM4EM, a novel approach for 3D segmentation of complex neural structures in electron microscopy (EM) data by leveraging the Segment Anything Model (SAM) alongside advanced fine-tuning strategies. Our contributions include the…

Computer Vision and Pattern Recognition · Computer Science 2025-05-01 Uzair Shah , Marco Agus , Daniya Boges , Vanessa Chiappini , Mahmood Alzubaidi , Jens Schneider , Markus Hadwiger , Pierre J. Magistretti , Mowafa Househ , Corrado Calı

Cataract remains a leading cause of visual impairment worldwide, and early detection from retinal imaging is critical for timely intervention. We present a deep learning pipeline for cataract classification using the Ocular Disease…

Image and Video Processing · Electrical Eng. & Systems 2025-09-30 MohammadReza Abbaszadeh Bavil Soflaei , Karim SamadZamini

Oral epithelial dysplasia (OED) is a premalignant histopathological diagnosis given to lesions of the oral cavity. Its grading suffers from significant inter-/intra- observer variability, and does not reliably predict malignancy…

Irreversible visual impairment is often caused by primary angle-closure glaucoma, which could be detected via Anterior Segment Optical Coherence Tomography (AS-OCT). In this paper, an automated system based on deep learning is presented for…

Computer Vision and Pattern Recognition · Computer Science 2019-02-12 Huazhu Fu , Yanwu Xu , Stephen Lin , Damon Wing Kee Wong , Mani Baskaran , Meenakshi Mahesh , Tin Aung , Jiang Liu

We introduce the Segment Anything (SA) project: a new task, model, and dataset for image segmentation. Using our efficient model in a data collection loop, we built the largest segmentation dataset to date (by far), with over 1 billion…

Computer Vision and Pattern Recognition · Computer Science 2023-04-06 Alexander Kirillov , Eric Mintun , Nikhila Ravi , Hanzi Mao , Chloe Rolland , Laura Gustafson , Tete Xiao , Spencer Whitehead , Alexander C. Berg , Wan-Yen Lo , Piotr Dollár , Ross Girshick

Ophthalmic microsurgery is known to be a challenging operation, which requires very precise and dexterous manipulation. Image guided robot-assisted surgery (RAS) is a promising solution that brings significant improvements in outcomes and…

Fundus photography has been routinely used to document the presence and severity of various retinal degenerative diseases such as age-related macula degeneration, glaucoma, and diabetic retinopathy, for which the fovea, optic disc (OD), and…

Image and Video Processing · Electrical Eng. & Systems 2022-03-02 Huaqing He , Li Lin , Zhiyuan Cai , Xiaoying Tang

Segment anything model (SAM) demonstrates strong generalization ability on natural image segmentation. However, its direct adaptation in medical image segmentation tasks shows significant performance drops. It also requires an excessive…

Computer Vision and Pattern Recognition · Computer Science 2024-12-19 Heng Guo , Jianfeng Zhang , Jiaxing Huang , Tony C. W. Mok , Dazhou Guo , Ke Yan , Le Lu , Dakai Jin , Minfeng Xu

Optical endomicroscopy (OEM) is an emerging technology platform with preclinical and clinical imaging applications. Pulmonary OEM via fibre bundles has the potential to provide in vivo, in situ molecular signatures of disease such as…

Computer Vision and Pattern Recognition · Computer Science 2018-08-29 Ahmed Karam Eldaly , Yoann Altmann , Antonios Perperidis , Nikola Krstajic , Tushar Choudhary , Kevin Dhaliwal , Stephen McLaughlin

Gliomas, the most prevalent primary brain tumors, require precise segmentation for diagnosis and treatment planning. However, this task poses significant challenges, particularly in the African population, were limited access to…

Image and Video Processing · Electrical Eng. & Systems 2023-12-20 Mohannad Barakat , Noha Magdy , Jjuuko George William , Ethel Phiri , Raymond Confidence , Dong Zhang , Udunna C Anazodo

In this paper, we explore the zero-shot capability of the Segment Anything Model (SAM) for food image segmentation. To address the lack of class-specific information in SAM-generated masks, we propose a novel framework, called FoodSAM. This…

Computer Vision and Pattern Recognition · Computer Science 2023-11-10 Xing Lan , Jiayi Lyu , Hanyu Jiang , Kun Dong , Zehai Niu , Yi Zhang , Jian Xue

Training large neural networks on large-scale datasets requires substantial computational resources, particularly for dense prediction tasks such as object detection. Although dataset distillation (DD) has been proposed to alleviate these…

Computer Vision and Pattern Recognition · Computer Science 2026-04-21 Salwa K. Al Khatib , Ahmed ElHagry , Shitong Shao , Zhiqiang Shen

TomoSAM has been developed to integrate the cutting-edge Segment Anything Model (SAM) into 3D Slicer, a highly capable software platform used for 3D image processing and visualization. SAM is a promptable deep learning model that is able to…

Computer Vision and Pattern Recognition · Computer Science 2023-06-16 Federico Semeraro , Alexandre Quintart , Sergio Fraile Izquierdo , Joseph C. Ferguson

Vision-language segmentation models such as SAM3 enable flexible, prompt-driven visual grounding, but inherit large, general-purpose text encoders originally designed for open-ended language understanding. In practice, segmentation prompts…

Artificial Intelligence · Computer Science 2026-02-13 Chengxi Zeng , Yuxuan Jiang , Ge Gao , Shuai Wang , Duolikun Danier , Bin Zhu , Stevan Rudinac , David Bull , Fan Zhang

The Segment Anything Model (SAM) and similar models build a family of promptable foundation models (FMs) for image and video segmentation. The object of interest is identified using prompts, such as bounding boxes or points. With these FMs…

Computer Vision and Pattern Recognition · Computer Science 2024-11-14 Caroline Magg , Hoel Kervadec , Clara I. Sánchez

Optical coherence tomography (OCT) is a non-invasive, micrometer-scale imaging modality that has become a clinical standard in ophthalmology. By raster-scanning the retina, sequential cross-sectional image slices are acquired to generate…

Image and Video Processing · Electrical Eng. & Systems 2024-02-20 Stefan Ploner , Jungeun Won , Julia Schottenhamml , Jessica Girgis , Kenneth Lam , Nadia Waheed , James Fujimoto , Andreas Maier

Detection of Barrett's esophagus (BE) at points of care outside the endoscopy suite may improve screening access and reduce esophageal adenocarcinoma mortality. Tethered capsule optical coherence tomography (OCT) can volumetrically image…

Segmentation in medical imaging is a critical component for the diagnosis, monitoring, and treatment of various diseases and medical conditions. Presently, the medical segmentation landscape is dominated by numerous specialized deep…