中文
相关论文

相关论文: Pairing-free Group-level Knowledge Distillation fo…

200 篇论文

In recent years, deep convolutional neural networks have made significant advances in pathology image segmentation. However, pathology image segmentation encounters with a dilemma in which the higher-performance networks generally require…

图像与视频处理 · 电气工程与系统科学 2021-11-15 Wenxuan Zou , Muyi Sun

Accurate prediction of protein-ligand binding affinity plays a pivotal role in accelerating the discovery of novel drugs and vaccines, particularly for gastrointestinal (GI) diseases such as gastric ulcers, Crohn's disease, and ulcerative…

机器学习 · 计算机科学 2025-11-11 Ziyang Gao , Annie Cheung , Yihao Ou

The performance of imaging techniques has an important influence on the clinical diagnostic strategy of colorectal cancer. Linked color imaging (LCI) by laser endoscopy is a recently developed techniques, and its advantage in improving the…

图像与视频处理 · 电气工程与系统科学 2018-08-01 Xinran Wei , Jiyang Xie , Wenrui He , Min Min , Zhanyu Ma , Jun Guo

Staining is essential in cell imaging and medical diagnostics but poses significant challenges, including high cost, time consumption, labor intensity, and irreversible tissue alterations. Recent advances in deep learning have enabled…

计算机视觉与模式识别 · 计算机科学 2025-04-15 Ziwang Xu , Lanqing Guo , Satoshi Tsutsui , Shuyan Zhang , Alex C. Kot , Bihan Wen

Very High Resolution (VHR) forest structure data at individual-tree scale is essential for carbon, biodiversity, and ecosystem monitoring. Still, airborne LiDAR remains costly and infrequent despite being the reference for forest structure…

计算机视觉与模式识别 · 计算机科学 2026-04-03 Taimur Khan , Hannes Feilhauer , Muhammad Jazib Zafar

The increased amount of multi-modal medical data has opened the opportunities to simultaneously process various modalities such as imaging and non-imaging data to gain a comprehensive insight into the disease prediction domain. Recent…

Gait recognition is an attractive biometric modality for long-range and contact-free identification, but high-performing gait models often rely on deep and computationally expensive architectures that are difficult to deploy in practice.…

计算机视觉与模式识别 · 计算机科学 2026-04-30 Yuqi Li , Qian Zhou , Huiran Duan , Jingjie Wang , Shunli Zhang , Chuanguang Yang , Guoying Zhao , Yingli Tian

Deep learning is a powerful tool for whole slide image (WSI) analysis. Typically, when performing supervised deep learning, a WSI is divided into small patches, trained and the outcomes are aggregated to estimate disease grade. However,…

计算机视觉与模式识别 · 计算机科学 2022-05-20 Yi Zheng , Rushin H. Gindra , Emily J. Green , Eric J. Burks , Margrit Betke , Jennifer E. Beane , Vijaya B. Kolachalama

Due to the surging amount of AI-generated images, its provisioning to edges and mobile users from the cloud incurs substantial traffic on networks. Generative semantic communication (GSC) offers a promising solution by transmitting highly…

机器学习 · 计算机科学 2026-01-27 Jingzhi Hu , Geoffrey Ye Li

Whole-slide images (WSIs) contain tissue information distributed across multiple magnification levels, yet most self-supervised methods treat these scales as independent views. This separation prevents models from learning representations…

Accurate segmentation of maxillary sinus in panoramic X-ray images is essential for dental diagnosis and surgical planning; however, this task remains relatively underexplored in dental imaging research. Structural overlap, ambiguous…

计算机视觉与模式识别 · 计算机科学 2026-04-23 Juha Park , Jiho Choi , Jong Pil Yun , Yong Chan Park , Han-Gyeol Yeom , Byung Do Lee , Sang Jun Lee

Monocular 3D object detection is an inherently ill-posed problem, as it is challenging to predict accurate 3D localization from a single image. Existing monocular 3D detection knowledge distillation methods usually project the LiDAR onto…

计算机视觉与模式识别 · 计算机科学 2023-10-18 Sen Wang , Jin Zheng

Multi-modal RGB and Depth (RGBD) data are predominant in many domains such as robotics, autonomous driving and remote sensing. The combination of these multi-modal data enhances environmental perception by providing 3D spatial context,…

计算机视觉与模式识别 · 计算机科学 2025-06-02 Roger Ferrod , Cássio F. Dantas , Luigi Di Caro , Dino Ienco

Hyperspectral imaging, capturing detailed spectral information for each pixel, is pivotal in diverse scientific and industrial applications. Yet, the acquisition of high-resolution (HR) hyperspectral images (HSIs) often needs to be…

计算机视觉与模式识别 · 计算机科学 2024-07-01 Chih-Chung Hsu , Chih-Chien Ni , Chia-Ming Lee , Li-Wei Kang

The widespread adoption of large-scale pre-training techniques has significantly advanced the development of medical foundation models, enabling them to serve as versatile tools across a broad range of medical tasks. However, despite their…

计算机视觉与模式识别 · 计算机科学 2024-10-22 Haolin Li , Yuhang Zhou , Ziheng Zhao , Siyuan Du , Jiangchao Yao , Weidi Xie , Ya Zhang , Yanfeng Wang

Disease classification relying solely on imaging data attracts great interest in medical image analysis. Current models could be further improved, however, by also employing Electronic Health Records (EHRs), which contain rich information…

图像与视频处理 · 电气工程与系统科学 2021-03-22 Tom van Sonsbeek , Xiantong Zhen , Marcel Worring , Ling Shao

Automated detection of Gallbladder Cancer (GBC) from Ultrasound (US) images is an important problem, which has drawn increased interest from researchers. However, most of these works use difficult-to-acquire information such as bounding box…

计算机视觉与模式识别 · 计算机科学 2023-09-12 Soumen Basu , Ashish Papanai , Mayank Gupta , Pankaj Gupta , Chetan Arora

Recent advances in knowledge distillation have emphasized the importance of decoupling different knowledge components. While existing methods utilize momentum mechanisms to separate task-oriented and distillation gradients, they overlook…

计算机视觉与模式识别 · 计算机科学 2025-05-22 Haiduo Huang , Jiangcheng Song , Yadong Zhang , Pengju Ren

Vision Transformers (ViTs) have achieved significant advancement in computer vision tasks due to their powerful modeling capacity. However, their performance notably degrades when trained with insufficient data due to lack of inherent…

图像与视频处理 · 电气工程与系统科学 2025-03-04 Omar S. EL-Assiouti , Ghada Hamed , Dina Khattab , Hala M. Ebied

Vision Transformers (ViTs) emerge to achieve impressive performance on many data-abundant computer vision tasks by capturing long-range dependencies among local features. However, under few-shot learning (FSL) settings on small datasets…

计算机视觉与模式识别 · 计算机科学 2023-03-31 Han Lin , Guangxing Han , Jiawei Ma , Shiyuan Huang , Xudong Lin , Shih-Fu Chang