中文
相关论文

相关论文: Context-Aware Optimal Transport Learning for Retin…

200 篇论文

Fundus images are very useful in identifying various ophthalmic disorders. However, due to the presence of artifacts, the visibility of the retina is severely affected. This may result in misdiagnosis of the disorder which may lead to more…

图像与视频处理 · 电气工程与系统科学 2021-12-28 Sai Koushik S S , K. G. Srinivasa

This thesis examines self-attention training through the lens of Optimal Transport (OT) and develops an OT-based alternative for tabular classification. The study tracks intermediate projections of the self-attention layer during training…

机器学习 · 统计学 2026-02-19 Alessandro Quadrio , Antonio Candelieri

Cataracts are the leading cause of vision loss worldwide. Restoration algorithms are developed to improve the readability of cataract fundus images in order to increase the certainty in diagnosis and treatment for cataract patients.…

图像与视频处理 · 电气工程与系统科学 2022-10-19 Heng Li , Haofeng Liu , Yan Hu , Huazhu Fu , Yitian Zhao , Hanpei Miao , Jiang Liu

Semi-supervised learning has made remarkable strides by effectively utilizing a limited amount of labeled data while capitalizing on the abundant information present in unlabeled data. However, current algorithms often prioritize aligning…

计算机视觉与模式识别 · 计算机科学 2024-05-31 Zhiquan Tan , Kaipeng Zheng , Weiran Huang

Although self-supervised learning enables us to bootstrap the training by exploiting unlabeled data, the generic self-supervised methods for natural images do not sufficiently incorporate the context. For medical images, a desirable method…

图像与视频处理 · 电气工程与系统科学 2022-07-08 Li Sun , Ke Yu , Kayhan Batmanghelich

Optimal transport is a machine learning problem with applications including distribution comparison, feature selection, and generative adversarial networks. In this paper, we propose feature-robust optimal transport (FROT) for…

We study the use of amortized optimization to predict optimal transport (OT) maps from the input measures, which we call Meta OT. This helps repeatedly solve similar OT problems between different measures by leveraging the knowledge and…

机器学习 · 计算机科学 2023-06-06 Brandon Amos , Samuel Cohen , Giulia Luise , Ievgen Redko

We introduce SceneTransporter, an end-to-end framework for structured 3D scene generation from a single image. While existing methods generate part-level 3D objects, they often fail to organize these parts into distinct instances in…

计算机视觉与模式识别 · 计算机科学 2026-02-27 Ling Wang , Hao-Xiang Guo , Xinzhou Wang , Fuchun Sun , Kai Sun , Pengkun Liu , Hang Xiao , Zhong Wang , Guangyuan Fu , Eric Li , Yang Liu , Yikai Wang

It is feasible to recognize the presence and seriousness of eye disease by investigating the progressions in retinal biological structure. Fundus examination is a diagnostic procedure to examine the biological structure and anomaly of the…

图像与视频处理 · 电气工程与系统科学 2022-07-19 Amit Bhati , Neha Gour , Pritee Khanna , Aparajita Ojha

Recent advances in deep learning, such as powerful generative models and joint text-image embeddings, have provided the computational creativity community with new tools, opening new perspectives for artistic pursuits. Text-to-image…

计算机视觉与模式识别 · 计算机科学 2022-04-20 Yingtao Tian , Marco Cuturi , David Ha

Diabetic retinopathy (DR) grading from fundus images has attracted increasing interest in both academic and industrial communities. Most convolutional neural network (CNN) based algorithms treat DR grading as a classification task via…

计算机视觉与模式识别 · 计算机科学 2021-03-19 Yehui Yang , Fangxin Shang , Binghong Wu , Dalu Yang , Lei Wang , Yanwu Xu , Wensheng Zhang , Tianzhu Zhang

Reliable traversable area segmentation in unstructured environments is critical for planning and decision-making in autonomous driving. However, existing data-driven approaches often suffer from degraded segmentation performance in…

计算机视觉与模式识别 · 计算机科学 2026-01-16 Zhihua Zhao , Guoqiang Li , Chen Min , Kangping Lu

We scrutinise an important observation plaguing scene-level sketch research -- that a significant portion of scene sketches are "partial". A quick pilot study reveals: (i) a scene sketch does not necessarily contain all objects in the…

计算机视觉与模式识别 · 计算机科学 2022-03-29 Pinaki Nath Chowdhury , Ayan Kumar Bhunia , Viswanatha Reddy Gajjala , Aneeshan Sain , Tao Xiang , Yi-Zhe Song

Cross-domain alignment between two sets of entities (e.g., objects in an image, words in a sentence) is fundamental to both computer vision and natural language processing. Existing methods mainly focus on designing advanced attention…

计算与语言 · 计算机科学 2020-07-28 Liqun Chen , Zhe Gan , Yu Cheng , Linjie Li , Lawrence Carin , Jingjing Liu

In contrast to non-medical image denoising, where enhancing image clarity is the primary goal, medical image denoising warrants preservation of crucial features without introduction of new artifacts. However, many denoising methods that…

图像与视频处理 · 电气工程与系统科学 2024-12-02 Md. Touhidul Islam , Md. Abtahi M. Chowdhury , Sumaiya Salekin , Aye T. Maung , Akil A. Taki , Hafiz Imtiaz

Optic nerve head (ONH) detection has been a crucial area of study in ophthalmology for years. However, the significant discrepancy between fundus image datasets, each generated using a single type of fundus camera, poses challenges to the…

图像与视频处理 · 电气工程与系统科学 2024-06-04 Jiayi Wang , Yi-An Mao , Xiaoyu Ma , Sicen Guo , Yuting Shao , Xiao Lv , Wenting Han , Mark Christopher , Linda M. Zangwill , Yanlong Bi , Rui Fan

The automatic diagnosis of various retinal diseases from fundus images is important to support clinical decision-making. However, developing such automatic solutions is challenging due to the requirement of a large amount of human-annotated…

计算机视觉与模式识别 · 计算机科学 2020-07-23 Xiaomeng Li , Mengyu Jia , Md Tauhidul Islam , Lequan Yu , Lei Xing

Ultra-Wide-Field (UWF) retinal imaging has revolutionized retinal diagnostics by providing a comprehensive view of the retina. However, it often suffers from quality-degrading factors such as blurring and uneven illumination, which obscure…

计算机视觉与模式识别 · 计算机科学 2025-08-28 Weicheng Liao , Zan Chen , Jianyang Xie , Yalin Zheng , Yuhui Ma , Yitian Zhao

Identifying lesions in fundus images is an important milestone toward an automated and interpretable diagnosis of retinal diseases. To support research in this direction, multiple datasets have been released, proposing groundtruth maps for…

计算机视觉与模式识别 · 计算机科学 2024-05-15 Clément Playout , Farida Cheriet

Multi-label image classification is a prediction task that aims to identify more than one label from a given image. This paper considers the semantic consistency of the latent space between the visual patch and linguistic label domains and…

计算机视觉与模式识别 · 计算机科学 2023-08-21 Miaoge Li , Dongsheng Wang , Xinyang Liu , Zequn Zeng , Ruiying Lu , Bo Chen , Mingyuan Zhou