中文
相关论文

相关论文: Adjustable Robust Transformer for High Myopia Scre…

200 篇论文

Optical coherence tomography (OCT) is a non-invasive volumetric imaging modality with high spatial and temporal resolution. For imaging larger tissue structures, OCT probes need to be moved to scan the respective area. For handheld…

计算机视觉与模式识别 · 计算机科学 2026-03-24 Suresh Guttikonda , Maximilian Neidhardt , Vidas Raudonis , Alexander Schlaefer

The trade-off between angular resolution and acceptance in scattering-angle measurements with a magnetic spectrometer is quantitatively evaluated for the Large Acceptance Spectrometer (LAS). The dependence on the multipole magnet field…

Imaging through complex scattering media is severely limited by aberrations and scattering which obscure images and reduce resolution. Confocal and temporal gatings partly filter out multiple scattering but are severely degraded by…

光学 · 物理学 2026-05-18 Yiwen Zhang , Minh Dinh , Zeyu Wang , Tianhao Zhang , Tianhang Chen , Chia Wei Hsu

Deep learning has substantially advanced pansharpening, achieving impressive fusion quality. However, a prevalent limitation is that conventional deep learning models, which typically rely on training datasets, often exhibit suboptimal…

计算机视觉与模式识别 · 计算机科学 2025-05-19 Haorui Chen , Zeyu Ren , Jiaxuan Ren , Ran Ran , Jinliang Shao , Jie Huang , Liangjian Deng

Rotary Positional Embeddings (RoPE) have become the standard for Large Language Models (LLMs) due to their ability to encode relative positions through geometric rotation. However, we identify a significant limitation we term ''Spectral…

计算与语言 · 计算机科学 2026-02-02 Kanishk Awadhiya

Homography estimation is a basic computer vision task, which aims to obtain the transformation from multi-view images for image alignment. Unsupervised learning homography estimation trains a convolution neural network for feature…

计算机视觉与模式识别 · 计算机科学 2023-02-07 Mingxiao Huo , Zhihao Zhang , Xinyang Ren , Xianqiang Yang

This paper discusses how ophthalmologists often rely on multimodal data to improve diagnostic accuracy. However, complete multimodal data is rare in real-world applications due to a lack of medical equipment and concerns about data privacy.…

计算机视觉与模式识别 · 计算机科学 2025-06-26 Xinkun Wang , Yifang Wang , Senwei Liang , Feilong Tang , Chengzhi Liu , Ming Hu , Chao Hu , Junjun He , Zongyuan Ge , Imran Razzak

With the commissioning of the refurbished adaptive secondary mirror (ASM) for the 6.5-meter MMT Observatory under way, special consideration had to be made to properly calibrate the mirror response functions to generate an interaction…

The scarcity of annotated data, particularly for rare diseases, limits the variability of training data and the range of detectable lesions, presenting a significant challenge for supervised anomaly detection in medical imaging. To solve…

计算机视觉与模式识别 · 计算机科学 2023-08-30 Yiming Huang , Guole Liu , Yaoru Luo , Ge Yang

Deep optics has emerged as a promising approach by co-designing optical elements with deep learning algorithms. However, current research typically overlooks the analysis and optimization of manufacturing and assembly tolerances. This…

计算机视觉与模式识别 · 计算机科学 2025-02-10 Jun Dai , Liqun Chen , Xinge Yang , Yuyao Hu , Jinwei Gu , Tianfan Xue

As new large-scale astronomical surveys greatly increase the number of objects targeted and discoveries made, the requirement for efficient follow-up observations is crucial. Adaptive optics imaging, which compensates for the image-blurring…

Aerial Remote Sensing (ARS) vision tasks present significant challenges due to the unique viewing angle characteristics. Existing research has primarily focused on algorithms for specific tasks, which have limited applicability in a broad…

计算机视觉与模式识别 · 计算机科学 2025-09-17 Wenhui Diao , Haichen Yu , Kaiyue Kang , Tong Ling , Di Liu , Yingchao Feng , Hanbo Bi , Libo Ren , Xuexue Li , Yongqiang Mao , Xian Sun

Quantifying the degree of atrophy is done clinically by neuroradiologists following established visual rating scales. For these assessments to be reliable the rater requires substantial training and experience, and even then the rating…

Early identification of stroke is crucial for intervention, requiring reliable models. We proposed an efficient retinal image representation together with clinical information to capture a comprehensive overview of cardiovascular health,…

计算机视觉与模式识别 · 计算机科学 2024-11-11 Yuqing Huang , Bastian Wittmann , Olga Demler , Bjoern Menze , Neda Davoudi

High quality structural volumetric imaging is a challenging goal to achieve with modern ultrasound transducers. Matrix probes have limited fields of view and element counts, whereas row-column arrays (RCAs) provide insufficient focusing. In…

图像与视频处理 · 电气工程与系统科学 2026-01-29 Darren Dahunsi , Randy Palamar , Tyler Henry , Mohammad Rahim Sobhani , Negar Majidi , Joy Wang , Afshin Kashani Ilkhechi , Roger Zemp

Stroke is a major public health problem, affecting millions worldwide. Deep learning has recently demonstrated promise for enhancing the diagnosis and risk prediction of stroke. However, existing methods rely on costly medical imaging…

图像与视频处理 · 电气工程与系统科学 2025-12-17 Saeed Shurrab , Aadim Nepal , Terrence J. Lee-St. John , Nicola G. Ghazi , Bartlomiej Piechowski-Jozwiak , Farah E. Shamout

We propose a novel framework for Alzheimer's disease (AD) detection using brain MRIs. The framework starts with a data augmentation method called Brain-Aware Replacements (BAR), which leverages a standard brain parcellation to replace…

计算机视觉与模式识别 · 计算机科学 2022-07-22 Mehmet Saygın Seyfioğlu , Zixuan Liu , Pranav Kamath , Sadjyot Gangolli , Sheng Wang , Thomas Grabowski , Linda Shapiro

Scene Text Recognition (STR) is challenging in extracting effective character representations from visual data when text is unreadable. Permutation language modeling (PLM) is introduced to refine character predictions by jointly capturing…

计算机视觉与模式识别 · 计算机科学 2026-02-04 Honghui Chen , Yuhang Qiu , Jiabao Wang , Pingping Chen , Nam Ling

Template matching is a fundamental task in computer vision and has been studied for decades. It plays an essential role in manufacturing industry for estimating the poses of different parts, facilitating downstream tasks such as robotic…

计算机视觉与模式识别 · 计算机科学 2024-08-21 Zhirui Gao , Renjiao Yi , Zheng Qin , Yunfan Ye , Chenyang Zhu , Kai Xu

Image-to-image translation is an ill-posed problem as unique one-to-one mapping may not exist between the source and target images. Learning-based methods proposed in this context often evaluate the performance on test data that is similar…

图像与视频处理 · 电气工程与系统科学 2021-10-08 Uddeshya Upadhyay , Viswanath P. Sudarshan , Suyash P. Awate