English
Related papers

Related papers: Self-Supervised Pretraining on Paired Sequences of…

200 papers

Accurate prediction of material properties facilitates the discovery of novel materials with tailored functionalities. Deep learning models have recently shown superior accuracy and flexibility in capturing structure-property relationships.…

Machine Learning · Computer Science 2025-04-30 Chowdhury Mohammad Abid Rahman , Aldo H. Romero , Prashnna K. Gyawali

Despite the impressive advances achieved using deep learning for functional brain activity analysis, the heterogeneity of functional patterns and the scarcity of imaging data still pose challenges in tasks such as identifying neurological…

Image and Video Processing · Electrical Eng. & Systems 2025-05-30 Wenhui Cui , Haleh Akrami , Anand A. Joshi , Richard M. Leahy

Pre-training lays the foundation for recent successes in radiograph analysis supported by deep learning. It learns transferable image representations by conducting large-scale fully-supervised or self-supervised learning on a source domain.…

Image and Video Processing · Electrical Eng. & Systems 2022-01-28 Hong-Yu Zhou , Xiaoyu Chen , Yinghao Zhang , Ruibang Luo , Liansheng Wang , Yizhou Yu

Deep learning associated with neurological signals is poised to drive major advancements in diverse fields such as medical diagnostics, neurorehabilitation, and brain-computer interfaces. The challenge in harnessing the full potential of…

Signal Processing · Electrical Eng. & Systems 2024-07-08 Di Wu , Siyuan Li , Jie Yang , Mohamad Sawan

We present Music Tagging Transformer that is trained with a semi-supervised approach. The proposed model captures local acoustic characteristics in shallow convolutional layers, then temporally summarizes the sequence of the extracted…

Sound · Computer Science 2021-11-29 Minz Won , Keunwoo Choi , Xavier Serra

Self-supervised learning has demonstrated considerable potential in hyperspectral representation, yet its application in cross-domain transfer scenarios remains under-explored. Existing methods, however, still rely on source domain…

Computer Vision and Pattern Recognition · Computer Science 2026-01-27 Jianshu Chao , Tianhua Lv , Qiqiong Ma , Yunfei Qiu , Li Fang , Huifang Shen , Wei Yao

Many current deep learning approaches make extensive use of backbone networks pre-trained on large datasets like ImageNet, which are then fine-tuned to perform a certain task. In remote sensing, the lack of comparable large annotated…

Computer Vision and Pattern Recognition · Computer Science 2024-08-22 Konrad Heidler , Lichao Mou , Di Hu , Pu Jin , Guangyao Li , Chuang Gan , Ji-Rong Wen , Xiao Xiang Zhu

We present a new self-supervised pre-training of Vision Transformers for dense prediction tasks. It is based on a contrastive loss across views that compares pixel-level representations to global image representations. This strategy…

Computer Vision and Pattern Recognition · Computer Science 2022-06-08 Jaonary Rabarisoa , Valentin Belissen , Florian Chabot , Quoc-Cuong Pham

\hspace{2mm} Diffusion-weighted magnetic resonance imaging (dMRI) of the brain offers unique capabilities including noninvasive probing of tissue microstructure and structural connectivity. It is widely used for clinical assessment of…

Image and Video Processing · Electrical Eng. & Systems 2025-12-30 Davood Karimi , Simon K. Warfield

In neuroscience, understanding inter-individual differences has recently emerged as a major challenge, for which functional magnetic resonance imaging (fMRI) has proven invaluable. For this, neuroscientists rely on basic methods such as…

Computer Vision and Pattern Recognition · Computer Science 2020-04-07 Akrem Sellami , François-Xavier Dupé , Bastien Cagna , Hachem Kadri , Stéphane Ayache , Thierry Artières , Sylvain Takerkart

We approached the goal of applying meta-learning to self-supervised masked autoencoders for spatiotemporal learning in three steps. Broadly, we seek to understand the impact of applying meta-learning to existing state-of-the-art…

Computer Vision and Pattern Recognition · Computer Science 2023-08-07 Faraz Waseem , Pratyush Muthukumar

Direct speech-to-text translation systems encounter an important drawback in data scarcity. A common solution consists on pretraining the encoder on automatic speech recognition, hence losing efficiency in the training process. In this…

Computation and Language · Computer Science 2024-09-27 Belen Alastruey , Gerard I. Gállego , Marta R. Costa-jussà

Clinical deployment of automated brain MRI analysis faces a fundamental challenge: clinical data is heterogeneous and noisy, and high-quality labels are prohibitively costly to obtain. Self-supervised learning (SSL) can address this by…

Computer Vision and Pattern Recognition · Computer Science 2026-05-25 Asbjørn Munk , Stefano Cerri , Vardan Nersesjan , Christian Hedeager Krag , Jakob Ambsdorf , Pablo Rocamora García , Julia Machnio , Peirong Liu , Suhyun Ahn , Nasrin Akbari , Yasmina Al Khalil , Kimberly Amador , Sina Amirrajab , Tal Arbel , Meritxell Bach Cuadra , Ujjwal Baid , Bhakti Baheti , Jaume Banus , Kamil Barbierik , Christoph Brune , Yansong Bu , Baptiste Callard , Yuhan Chen , Cornelius Crijnen , Corentin Dancette , Peter Drotar , Prasad Dutande , Nils D. Forkert , Saurabh Garg , Jakub Gazda , Matej Gazda , Benoît Gérin , Partha Ghosh , Weikang Gong , Pedro M. Gordaliza , Sam Hashemi , Tobias Heimann , Fucang Jia , Jiexin Jiang , Emily Kaczmarek , Chris Kang , Seung Kwan Kang , Mohammad Khazaei , Julien Khlaut , Petros Koutsouvelis , Jae Sung Lee , Yuchong Li , Mengye Lyu , Mingchen Ma , Anant Madabhushi , Klaus H. Maier-Hein , Pierre Manceron , Andrés Martínez Mora , Moona Mazher , Felix Meister , Nataliia Molchanova , Steven A. Niederer , Leonard Nürnberg , Jinah Park , Abdul Qayyum , Jonas Richiardi , Antoine Saporta , Branislav Setlak , Ning Shen , Justin Szeto , Constantin Ulrich , Puru Vaish , Vibujithan Vigneshwaran , Leroy Volmer , Zihao Wang , Siqi Wei , Anthony Winder , Jelmer M. Wolterink , Maxence Wynen , Chang Yang , Si Young Yie , Mostafa Mehdipour Ghazi , Akshay Pai , Espen Jimenez Solem , Sebastian Nørgaard Llambias , Mikael Boesen , Michael Eriksen Benros , Juan Eugenio Iglesias , Mads Nielsen

Self-supervised learning, which benefits from automatically constructing labels through pre-designed pretext task, has recently been applied for strengthen supervised learning. Since previous self-supervised pretext tasks are based on…

Computer Vision and Pattern Recognition · Computer Science 2021-06-10 Zilin Ding , Yuhang Yang , Xuan Cheng , Xiaomin Wang , Ming Liu

Speech enhancement (SE) is usually required as a front end to improve the speech quality in noisy environments, while the enhanced speech might not be optimal for automatic speech recognition (ASR) systems due to speech distortion. On the…

Audio and Speech Processing · Electrical Eng. & Systems 2022-05-27 Qiu-Shi Zhu , Jie Zhang , Zi-Qiang Zhang , Li-Rong Dai

Multi-pitch estimation is a decades-long research problem involving the detection of pitch activity associated with concurrent musical events within multi-instrument mixtures. Supervised learning techniques have demonstrated solid…

Audio and Speech Processing · Electrical Eng. & Systems 2024-02-27 Frank Cwitkowitz , Zhiyao Duan

Foundation models have reshaped the landscape of Remote Sensing (RS) by enhancing various image interpretation tasks. Pretraining is an active research topic, encompassing supervised and self-supervised learning methods to initialize model…

Computer Vision and Pattern Recognition · Computer Science 2024-05-31 Di Wang , Jing Zhang , Minqiang Xu , Lin Liu , Dongsheng Wang , Erzhong Gao , Chengxi Han , Haonan Guo , Bo Du , Dacheng Tao , Liangpei Zhang

Accurate classification of medical device risk levels is essential for regulatory oversight and clinical safety. We present a Transformer-based multimodal framework that integrates textual descriptions and visual information to predict…

Machine Learning · Computer Science 2025-05-02 Yu Han , Aaron Ceross , Jeroen H. M. Bergmann

In network representation learning we learn how to represent heterogeneous information networks in a low-dimensional space so as to facilitate effective search, classification, and prediction solutions. Previous network representation…

Artificial Intelligence · Computer Science 2021-05-19 Yang Fang , Xiang Zhao , Yifan Chen , Weidong Xiao , Maarten de Rijke

In this paper, we show that a simple self-supervised pre-trained audio model can achieve comparable inference efficiency to more complicated pre-trained models with speech transformer encoders. These speech transformers rely on mixing…

Sound · Computer Science 2024-02-09 Sungho Jeon , Ching-Feng Yeh , Hakan Inan , Wei-Ning Hsu , Rashi Rungta , Yashar Mehdad , Daniel Bikel
‹ Prev 1 8 9 10 Next ›