English
Related papers

Related papers: Semantic-aware Temporal Channel-wise Attention for…

200 papers

Characteristics such as low contrast and significant organ shape variations are often exhibited in medical images. The improvement of segmentation performance in medical imaging is limited by the generally insufficient adaptive capabilities…

Image and Video Processing · Electrical Eng. & Systems 2023-06-09 Hejun Huang , Zuguo Chen , Ying Zou , Ming Lu , Chaoyang Chen

Deep learning organ segmentation approaches require large amounts of annotated training data, which is limited in supply due to reasons of confidentiality and the time required for expert manual annotation. Therefore, being able to train…

Computer Vision and Pattern Recognition · Computer Science 2020-08-14 Abdelrahman Elskhawy , Aneta Lisowska , Matthias Keicher , Josep Henry , Paul Thomson , Nassir Navab

Differential medical VQA models compare multiple images to identify clinically meaningful changes and rely on vision encoders to capture fine-grained visual differences that reflect radiologists' comparative diagnostic workflows. However,…

Computer Vision and Pattern Recognition · Computer Science 2026-04-23 Denis Musinguzi , Caren Han , Prasenjit Mitra

Long video understanding is a complex task that requires both spatial detail and temporal awareness. While Vision-Language Models (VLMs) obtain frame-level understanding capabilities through multi-frame input, they suffer from information…

Computer Vision and Pattern Recognition · Computer Science 2025-04-10 Ziyi Wang , Haoran Wu , Yiming Rong , Deyang Jiang , Yixin Zhang , Yunlong Zhao , Shuang Xu , Bo XU

Ejection fraction (EF) is commonly measured by echocardiography, by dividing the volume ejected by the heart (stroke volume) by the volume of the filled heart (end-diastolic volume). Utilizing volume changes of left myocardial segments per…

Medical Physics · Physics 2018-03-09 Mersedeh Karvandi , Saeed Ranjbar

Cardiotocography (CTG) is the main tool used for fetal monitoring during labour. Interpretation of CTG requires dynamic pattern recognition in real time. It is recognised as a difficult task with high inter- and intra-observer disagreement.…

Machine Learning · Computer Science 2021-11-02 M. O'Sullivan , T. Gabruseva , G. Boylan , M. O'Riordan , G. Lightbody , W. Marnane

Recently, learned video compression has achieved exciting performance. Following the traditional hybrid prediction coding framework, most learned methods generally adopt the motion estimation motion compensation (MEMC) method to remove…

Image and Video Processing · Electrical Eng. & Systems 2023-10-20 Yiming Wang , Qian Huang , Bin Tang , Huashan Sun , Xing Li

Electrocardiograms (ECGs) are among the most widely available clinical signals and play a central role in cardiovascular diagnosis. While recent foundation models (FMs) have shown promise for learning transferable ECG representations, most…

Echocardiography is widely used for assessing cardiac function, where clinically meaningful parameters such as left-ventricular ejection fraction (EF) play a central role in diagnosis and management. Generative models capable of…

Image and Video Processing · Electrical Eng. & Systems 2026-03-17 Emmanuel Oladokun , Sarina Thomas , Jurica Šprem , Vicente Grau

Segmentation of cardiac structures is one of the fundamental steps to estimate volumetric indices of the heart. This step is still performed semi-automatically in clinical routine, and is thus prone to inter- and intra-observer variability.…

Convolutional neural networks (CNN) have demonstrated their ability to segment 2D cardiac ultrasound images. However, despite recent successes according to which the intra-observer variability on end-diastole and end-systole images has been…

Image and Video Processing · Electrical Eng. & Systems 2022-05-09 Nathan Painchaud , Nicolas Duchateau , Olivier Bernard , Pierre-Marc Jodoin

Cardiac left ventricle (LV) quantification is among the most clinically important tasks for identification and diagnosis of cardiac diseases, yet still a challenge due to the high variability of cardiac structure and the complexity of…

Computer Vision and Pattern Recognition · Computer Science 2017-06-15 Wufeng Xue , Andrea Lum , Ashley Mercado , Mark Landis , James Warringto , Shuo Li

We present a new architecture for human action forecasting from videos. A temporal recurrent encoder captures temporal information of input videos while a self-attention model is used to attend on relevant feature dimensions of the input…

Computer Vision and Pattern Recognition · Computer Science 2021-07-20 Yan Bin Ng , Basura Fernando

This paper proposes a method for long-term action anticipation (LTA), the task of predicting action labels and their duration in a video given the observation of an initial untrimmed video interval. We build on an encoder-decoder…

Computer Vision and Pattern Recognition · Computer Science 2024-12-30 Alberto Maté , Mariella Dimiccoli

The temporal segmentation of events is an essential task and a precursor for the automatic recognition of human actions in the video. Several attempts have been made to capture frame-level salient aspects through attention but they lack the…

Computer Vision and Pattern Recognition · Computer Science 2020-05-08 Harshala Gammulle , Simon Denman , Sridha Sridharan , Clinton Fookes

Action visual tempo characterizes the dynamics and the temporal scale of an action, which is helpful to distinguish human actions that share high similarities in visual dynamics and appearance. Previous methods capture the visual tempo…

Computer Vision and Pattern Recognition · Computer Science 2022-07-13 Yuanzhong Liu , Junsong Yuan , Zhigang Tu

Semantic segmentation using convolutional neural networks (CNNs) is the state-of-the-art for many medical image segmentation tasks including myocardial segmentation in cardiac MR images. However, the predicted segmentation maps obtained…

Image and Video Processing · Electrical Eng. & Systems 2022-08-18 Sofie Tilborghs , Jan Bogaert , Frederik Maes

Owing to recent advances in thoracic electrical impedance tomography, a patient's hemodynamic function can be noninvasively and continuously estimated in real-time by surveilling a cardiac volume signal associated with stroke volume and…

Signal Processing · Electrical Eng. & Systems 2023-01-05 Chang Min Hyun , Tae Jun Jang , Jeongchan Nam , Hyeuknam Kwon , Kiwan Jeon , Kyunghun Lee

A major challenge for video captioning is to combine audio and visual cues. Existing multi-modal fusion methods have shown encouraging results in video understanding. However, the temporal structures of multiple modalities at different…

Computation and Language · Computer Science 2018-04-17 Xin Wang , Yuan-Fang Wang , William Yang Wang

The affect embedded in video data conveys high-level semantic information about the content and has direct impact on the understanding and perception of reviewers, as well as their emotional responses. Affective Video Content Analysis…

Multimedia · Computer Science 2018-06-04 Ligang Zhang , Jiulong Zhang
‹ Prev 1 4 5 6 7 8 10 Next ›