English
Related papers

Related papers: Seeing More with Less: Video Capsule Endoscopy wit…

200 papers

The accurate classification of gastrointestinal diseases from endoscopic and histopathological imagery remains a significant challenge in medical diagnostics, mainly due to the vast data volume and subtle variation in inter-class visuals.…

Image and Video Processing · Electrical Eng. & Systems 2026-04-28 Md Assaduzzaman , Nushrat Jahan Oyshi , Eram Mahamud

Multi-task learning (MTL) is a powerful approach in deep learning that leverages the information from multiple tasks during training to improve model performance. In medical imaging, MTL has shown great potential to solve various tasks.…

Computer Vision and Pattern Recognition · Computer Science 2023-09-08 Sangwook Kim , Thomas G. Purdie , Chris McIntosh

Capsule endoscopy (CE) enables non-invasive gastrointestinal screening, but current CE research remains largely limited to frame-level classification and detection, leaving video-level analysis underexplored. To bridge this gap, we…

Computer Vision and Pattern Recognition · Computer Science 2026-04-24 Bowen Liu , Li Yang , Shanshan Song , Mingyu Tang , Zhifang Gao , Qifeng Chen , Yangqiu Song , Huimin Chen , Xiaomeng Li

Video instance segmentation (VIS) is a new and critical task in computer vision. To date, top-performing VIS methods extend the two-stage Mask R-CNN by adding a tracking branch, leaving plenty of room for improvement. In contrast, we…

Computer Vision and Pattern Recognition · Computer Science 2021-04-06 Dongfang Liu , Yiming Cui , Wenbo Tan , Yingjie Chen

We aim to train a multi-task model such that users can adjust the desired compute budget and relative importance of task performances after deployment, without retraining. This enables optimizing performance for dynamically varying user…

Computer Vision and Pattern Recognition · Computer Science 2023-08-24 Abhishek Aich , Samuel Schulter , Amit K. Roy-Chowdhury , Manmohan Chandraker , Yumin Suh

Video streams are utilised to guide minimally-invasive surgery and diagnostic procedures in a wide range of procedures, and many computer assisted techniques have been developed to automatically analyse them. These approaches can provide…

Computer Vision and Pattern Recognition · Computer Science 2022-04-01 Rema Daher , Francisco Vasconcelos , Danail Stoyanov

Deep learning methods have surpassed the performance of traditional techniques on a wide range of problems in computer vision, but nearly all of this work has studied consumer photos, where precisely correct output is often not critical. It…

Computer Vision and Pattern Recognition · Computer Science 2018-07-24 Mingze Xu , Chenyou Fan , John D Paden , Geoffrey C Fox , David J Crandall

Multi-Task Learning (MTL) involves the concurrent training of multiple tasks, offering notable advantages for dense prediction tasks in computer vision. MTL not only reduces training and inference time as opposed to having multiple…

Computer Vision and Pattern Recognition · Computer Science 2024-12-05 Maxime Fontana , Michael Spratling , Miaojing Shi

Malnutrition is a multidomain problem affecting 54% of older adults in long-term care (LTC). Monitoring nutritional intake in LTC is laborious and subjective, limiting clinical inference capabilities. Recent advances in automatic…

Computer Vision and Pattern Recognition · Computer Science 2022-02-02 Kaylen J Pfisterer , Robert Amelard , Audrey G Chung , Braeden Syrnyk , Alexander MacLean , Heather H Keller , Alexander Wong

The lack of fine-grained annotations hinders the deployment of automated diagnosis systems, which require human-interpretable justification for their decision process. In this paper, we address the problem of weakly supervised…

Computer Vision and Pattern Recognition · Computer Science 2022-10-10 Constantin Seibold , Jens Kleesiek , Heinz-Peter Schlemmer , Rainer Stiefelhagen

Diabetic Retinopathy (DR) is a non-negligible eye disease among patients with Diabetes Mellitus, and automatic retinal image analysis algorithm for the DR screening is in high demand. Considering the resolution of retinal image is very…

Computer Vision and Pattern Recognition · Computer Science 2019-11-14 Kang Zhou , Zaiwang Gu , Wen Liu , Weixin Luo , Jun Cheng , Shenghua Gao , Jiang Liu

This work delves into unsupervised monocular depth estimation in endoscopy, which leverages adjacent frames to establish a supervisory signal during the training phase. For many clinical applications, e.g., surgical navigation, temporally…

Computer Vision and Pattern Recognition · Computer Science 2025-02-18 Shuwei Shao , Zhongcai Pei , Weihai Chen , Xingming Wu , Zhong Liu

The gastrointestinal (GI) tract of humans can have a wide variety of aberrant mucosal abnormality findings, ranging from mild irritations to extremely fatal illnesses. Prompt identification of gastrointestinal disorders greatly contributes…

Image and Video Processing · Electrical Eng. & Systems 2025-10-01 Sumaiya Tabassum , Md. Faysal Ahamed , Hafsa Binte Kibria , Md. Nahiduzzaman , Julfikar Haider , Muhammad E. H. Chowdhury , Mohammad Tariqul Islam

Nasogastric tubes (NGTs) are feeding tubes that are inserted through the nose into the stomach to deliver nutrition or medication. If not placed correctly, they can cause serious harm, even death to patients. Recent AI developments…

Traditional deep learning methods in medical imaging often focus solely on segmentation or classification, limiting their ability to leverage shared information. Multi-task learning (MTL) addresses this by combining both tasks through…

Image and Video Processing · Electrical Eng. & Systems 2024-12-03 Phuoc-Nguyen Bui , Duc-Tai Le , Junghyun Bum , Hyunseung Choo

Food volume estimation is an essential step in the pipeline of dietary assessment and demands the precise depth estimation of the food surface and table plane. Existing methods based on computer vision require either multi-image input or…

Computer Vision and Pattern Recognition · Computer Science 2020-08-04 Ya Lu , Thomai Stathopoulou , Stavroula Mougiakakou

Localization is an indispensable component of a robot's autonomy stack that enables it to determine where it is in the environment, essentially making it a precursor for any action execution or planning. Although convolutional neural…

Robotics · Computer Science 2018-03-13 Abhinav Valada , Noha Radwan , Wolfram Burgard

In this study, we introduce an innovative EEG signal reconstruction sub-module designed to enhance the performance of deep learning models on EEG eye-tracking tasks. This sub-module can integrate with all Encoder-Classifier-based deep…

Human-Computer Interaction · Computer Science 2024-08-13 Weigeng Li , Neng Zhou , Xiaodong Qu

Heterogeneous morphological features and data imbalance pose significant challenges in rare thyroid carcinoma classification using ultrasound imaging. To address this issue, we propose a novel multitask learning framework, Channel-Spatial…

Image and Video Processing · Electrical Eng. & Systems 2026-03-05 Peiqi Li , Yincheng Gao , Renxing Li , Haojie Yang , Yunyun Liu , Boji Liu , Jiahui Ni , Ying Zhang , Yulu Wu , Xiaowei Fang , Lehang Guo , Liping Sun , Jiangang Chen

Using lightweight models as backbone networks in gaze estimation tasks often results in significant performance degradation. The main reason is that the number of feature channels in lightweight networks is usually small, which makes the…

Computer Vision and Pattern Recognition · Computer Science 2024-12-10 Zhang Cheng , Yanxia Wang
‹ Prev 1 8 9 10 Next ›