中文
相关论文

相关论文: ResNet-50 with Class Reweighting and Anatomy-Guide…

200 篇论文

This work presents a multi-label temporal event detection framework for video capsule endoscopy (VCE) that addresses the extreme class imbalance inherent in the Galar dataset by combining two principal contributions: an Angular Separation…

计算机视觉与模式识别 · 计算机科学 2026-05-25 Podakanti Satyajith Chary , Nagarajan Ganapathy

Video Capsule Endoscopy (VCE) poses a challenging multi-label temporal classification problem, requiring simultaneous localization of 8 anatomical regions and detection of 9 pathological findings across tens of thousands of frames. We…

计算机视觉与模式识别 · 计算机科学 2026-05-22 Jiye Won , Seangmin Lee , Soon Ki Jung

This paper presents a deep learning framework for the multi-class classification of gastrointestinal abnormalities in Video Capsule Endoscopy (VCE) frames. The aim is to automate the identification of ten GI abnormality classes, including…

计算机视觉与模式识别 · 计算机科学 2024-10-25 Aman Sagar , Preeti Mehta , Monika Shrivastva , Suchi Kumari

Background: Twelve lead ECGs are a core diagnostic tool for cardiovascular diseases. Here, we describe and analyse an ensemble deep neural network architecture to classify 24 cardiac abnormalities from 12-lead ECGs. Method: We proposed a…

信号处理 · 电气工程与系统科学 2022-04-13 Zhibin Zhao , Darcy Murphy , Hugh Gifford , Stefan Williams , Annie Darlington , Samuel D. Relton , Hui Fang , David C. Wong

We propose a novel pathology-sensitive deep learning model (PS-DeVCEM) for frame-level anomaly detection and multi-label classification of different colon diseases in video capsule endoscopy (VCE) data. Our proposed model is capable of…

计算机视觉与模式识别 · 计算机科学 2020-11-30 A. Mohammed , I. Farup , M. Pedersen , S. Yildirim , Ø Hovde

This work is corresponding to the Gastro Competition for multi-label classification from capsule endoscopic videos (CEV). Deep learning network based on Transformers are fined-tune for this task. The based online mode is Google Vision…

计算机视觉与模式识别 · 计算机科学 2026-03-20 X. Gao , C. Chien , G. Liu , A. Manullang

To reveal the importance of temporal precision in ground truth audio event labels, we collected precise (~0.1 sec resolution) "strong" labels for a portion of the AudioSet dataset. We devised a temporally strong evaluation set (including…

We study multilabel classification of chest X-rays and present a simple, strong pipeline built on SE-ResNeXt101 $(32 \times 4d)$. The backbone is finetuned for 14 thoracic findings with a sigmoid head, trained using Multilabel Iterative…

计算机视觉与模式识别 · 计算机科学 2025-11-12 Utkarsh Prakash Srivastava , Kaushik Gupta , Kaushik Nath

Video-text retrieval has witnessed remarkable progress driven by large-scale vision-language pretraining, yet most existing approaches inherit an implicit assumption from image-text retrieval: that visual semantics can be captured…

计算机视觉与模式识别 · 计算机科学 2026-05-26 Zixu Li , Yupeng Hu , Zhiwei Chen , Zhiheng Fu , Xiaowei Zhu , Weili Guan , Liqiang Nie

We propose a novel method for real-time face alignment in videos based on a recurrent encoder-decoder network model. Our proposed model predicts 2D facial point heat maps regularized by both detection and regression loss, while uniquely…

计算机视觉与模式识别 · 计算机科学 2018-01-19 Xi Peng , Rogerio S. Feris , Xiaoyu Wang , Dimitris N. Metaxas

ImageNet has been arguably the most popular image classification benchmark, but it is also the one with a significant level of label noise. Recent studies have shown that many samples contain multiple classes, despite being assumed to be a…

计算机视觉与模式识别 · 计算机科学 2021-07-23 Sangdoo Yun , Seong Joon Oh , Byeongho Heo , Dongyoon Han , Junsuk Choe , Sanghyuk Chun

Capsule endoscopy event detection is challenging because diagnostically relevant findings are sparse, visually heterogeneous, and embedded in long, noisy video streams, while evaluation is performed at the event level rather than by frame…

计算机视觉与模式识别 · 计算机科学 2026-03-20 Bo-Cheng Qiu , Yu-Fan Lin , Yu-Zhe Pien , Chia-Ming Lee , Fu-En Yang , Yu-Chiang Frank Wang , Chih-Chung Hsu

For early diagnosis of malignancies in the gastrointestinal tract, surveillance endoscopy is increasingly used to monitor abnormal tissue changes in serial examinations of the same patient. Despite successes with optical biopsy for in vivo…

计算机视觉与模式识别 · 计算机科学 2016-11-01 Menglong Ye , Edward Johns , Benjamin Walter , Alexander Meining , Guang-Zhong Yang

Video capsule endoscopy has become increasingly important for investigating the small intestine within the gastrointestinal tract. However, a persistent challenge remains the short battery lifetime of such compact sensor edge devices.…

计算机视觉与模式识别 · 计算机科学 2026-01-21 Julia Werner , Oliver Bause , Julius Oexle , Maxime Le Floch , Franz Brinkmann , Jochen Hampe , Oliver Bringmann

Semantic video segmentation is challenging due to the sheer amount of data that needs to be processed and labeled in order to construct accurate models. In this paper we present a deep, end-to-end trainable methodology to video segmentation…

计算机视觉与模式识别 · 计算机科学 2017-10-03 David Nilsson , Cristian Sminchisescu

Fourteen million colonoscopies are performed annually just in the U.S. However, the videos from these colonoscopies are not saved due to storage constraints (each video from a high-definition colonoscope camera can be in tens of gigabytes).…

计算机视觉与模式识别 · 计算机科学 2025-05-14 Shawn Mathew , Saad Nadeem , Alvin C. Goh , Arie Kaufman

This paper introduces a novel segmentation framework that integrates a classifier network with a reverse HRNet architecture for efficient image segmentation. Our approach utilizes a ResNet-50 backbone, pretrained in a semi-supervised…

计算机视觉与模式识别 · 计算机科学 2024-02-12 Anupam Gupta , Ashok Krishnamurthy , Lisa Singh

Abnormalities in the gastrointestinal tract significantly influence the patient's health and require a timely diagnosis for effective treatment. With such consideration, an effective automatic classification of these abnormalities from a…

计算机视觉与模式识别 · 计算机科学 2024-12-24 Lakshmi Srinivas Panchananam , Praveen Kumar Chandaliya , Kishor Upla , Kiran Raja

Food image segmentation is a critical task for dietary analysis, enabling accurate estimation of food volume and nutrients. However, current methods suffer from limited multi-view data and poor generalization to new viewpoints. We introduce…

The increased availability of X-ray image archives (e.g. the ChestX-ray14 dataset from the NIH Clinical Center) has triggered a growing interest in deep learning techniques. To provide better insight into the different approaches, and their…

计算机视觉与模式识别 · 计算机科学 2019-01-30 Ivo M. Baltruschat , Hannes Nickisch , Michael Grass , Tobias Knopp , Axel Saalbach
‹ 上一页 1 2 3 10 下一页 ›