English
Related papers

Related papers: CTSL: Codebook-based Temporal-Spatial Learning for…

200 papers

Deep learning (DL) is a powerful tool in computational imaging for many applications. A common strategy is to reconstruct a preliminary image as the input of a neural network to achieve an optimized image. Usually, the preliminary image is…

Image and Video Processing · Electrical Eng. & Systems 2021-05-12 Ruibo Shang , Kevin Hoffer-Hawlik , Geoffrey P. Luke

Magnetic resonance imaging (MRI) is indispensable for diagnosing and planning treatment in various medical conditions due to its ability to produce multi-series images that reveal different tissue characteristics. However, integrating these…

Image and Video Processing · Electrical Eng. & Systems 2024-12-11 Churan Wang , Fei Gao , Lijun Yan , Siwen Wang , Yizhou Yu , Yizhou Wang

Self-supervised pretraining (SSP) has emerged as a popular technique in machine learning, enabling the extraction of meaningful feature representations without labelled data. In the realm of computer vision, pretrained vision transformers…

Computer Vision and Pattern Recognition · Computer Science 2023-08-23 Jiantao Wu , Shentong Mo , Muhammad Awais , Sara Atito , Zhenhua Feng , Josef Kittler

In recent years, the introduction of self-supervised contrastive learning (SSCL) has demonstrated remarkable improvements in representation learning across various domains, including natural language processing and computer vision. By…

Machine Learning · Computer Science 2023-08-15 Chiyu Zhang , Qi Yan , Lili Meng , Tristan Sylvain

Anomaly detection in multi-variate time series (MVTS) data is a huge challenge as it requires simultaneous representation of long term temporal dependencies and correlations across multiple variables. More often, this is solved by breaking…

Machine Learning · Computer Science 2022-02-09 Theivendiram Pranavan , Terence Sim , Arulmurugan Ambikapathi , Savitha Ramasamy

Crime has become a major concern in many cities, which calls for the rising demand for timely predicting citywide crime occurrence. Accurate crime prediction results are vital for the beforehand decision-making of government to alleviate…

Machine Learning · Computer Science 2022-08-19 Zhonghang Li , Chao Huang , Lianghao Xia , Yong Xu , Jian Pei

One main challenge in time series anomaly detection (TSAD) is the lack of labelled data in many real-life scenarios. Most of the existing anomaly detection methods focus on learning the normal behaviour of unlabelled time series in an…

Machine Learning · Computer Science 2024-09-04 Zahra Zamanzadeh Darban , Geoffrey I. Webb , Shirui Pan , Charu C. Aggarwal , Mahsa Salehi

Purpose: To develop a deep learning method on a nonlinear manifold to explore the temporal redundancy of dynamic signals to reconstruct cardiac MRI data from highly undersampled measurements. Methods: Cardiac MR image reconstruction is…

Image and Video Processing · Electrical Eng. & Systems 2021-04-05 Ziwen Ke , Zhuo-Xu Cui , Wenqi Huang , Jing Cheng , Sen Jia , Haifeng Wang , Xin Liu , Hairong Zheng , Leslie Ying , Yanjie Zhu , Dong Liang

To equip artificial intelligence with a comprehensive understanding towards a temporal world, video and 4D panoptic scene graph generation abstracts visual data into nodes to represent entities and edges to capture temporal relations.…

Computer Vision and Pattern Recognition · Computer Science 2026-04-28 Thong Thanh Nguyen , Xiaobao Wu , Yi Bin , Cong-Duy T Nguyen , See-Kiong Ng , Anh Tuan Luu

In this work, we present Multi-Level Contrastive Learning for Dense Prediction Task (MCL), an efficient self-supervised method for learning region-level feature representation for dense prediction tasks. Our method is motivated by the three…

Computer Vision and Pattern Recognition · Computer Science 2023-04-05 Qiushan Guo , Yizhou Yu , Yi Jiang , Jiannan Wu , Zehuan Yuan , Ping Luo

Continuous sign language recognition (CSLR) requires precise spatio-temporal modeling to accurately recognize sequences of gestures in videos. Existing frameworks often rely on CNN-based spatial backbones combined with temporal convolution…

Computer Vision and Pattern Recognition · Computer Science 2025-12-30 Ahmed Abul Hasanaath , Hamzah Luqman

Quantitative assessment of cardiac left ventricle (LV) morphology is essential to assess cardiac function and improve the diagnosis of different cardiovascular diseases. In current clinical practice, LV quantification depends on the…

Image and Video Processing · Electrical Eng. & Systems 2020-12-25 Sulaiman Vesal , Mingxuan Gu , Andreas Maier , Nishant Ravikumar

There are a number of studies about extraction of bottleneck (BN) features from deep neural networks (DNNs)trained to discriminate speakers, pass-phrases and triphone states for improving the performance of text-dependent speaker…

Sound · Computer Science 2019-05-14 Achintya kr. Sarkar , Zheng-Hua Tan , Hao Tang , Suwon Shon , James Glass

Motivation: CMR is the golden standard for cardiac diagnosis, and medical data annotation is time-consuming. Thus, screening techniques from unlabeled data can help streamline the cardiac diagnosis process. Goal: This work aims to enable…

Tissues and Organs · Quantitative Biology 2025-11-11 Yundi Zhang , Daniel Rueckert , Jiazhen Pan

The healthcare industry generates troves of unlabelled physiological data. This data can be exploited via contrastive learning, a self-supervised pre-training method that encourages representations of instances to be similar to one another.…

Machine Learning · Computer Science 2021-05-18 Dani Kiyasseh , Tingting Zhu , David A. Clifton

Semantic segmentation using convolutional neural networks (CNNs) is the state-of-the-art for many medical image segmentation tasks including myocardial segmentation in cardiac MR images. However, the predicted segmentation maps obtained…

Image and Video Processing · Electrical Eng. & Systems 2022-08-18 Sofie Tilborghs , Jan Bogaert , Frederik Maes

Although supervised learning has enabled high performance for image segmentation, it requires a large amount of labeled training data, which can be difficult to obtain in the medical imaging field. Self-supervised learning (SSL) methods…

Vision-language models (VLMs) such as CLIP have demonstrated remarkable zero-shot generalization, yet remain highly vulnerable to adversarial examples (AEs). While test-time defenses are promising, existing methods fail to provide…

Computer Vision and Pattern Recognition · Computer Science 2026-01-28 Sen Nie , Jie Zhang , Zhuo Wang , Shiguang Shan , Xilin Chen

With the rapid development of digital multimedia, video understanding has become an important field. For action recognition, temporal dimension plays an important role, and this is quite different from image recognition. In order to learn…

Computer Vision and Pattern Recognition · Computer Science 2020-02-11 Qian Liu , Tao Wang , Jie Liu , Yang Guan , Qi Bu , Longfei Yang

Dynamic MR images possess various transformation symmetries,including the rotation symmetry of local features within the image and along the temporal dimension. Utilizing these symmetries as prior knowledge can facilitate dynamic MR imaging…

Image and Video Processing · Electrical Eng. & Systems 2024-09-16 Yuliang Zhu , Jing Cheng , Zhuo-Xu Cui , Jianfeng Ren , Chengbo Wang , Dong Liang