中文
相关论文

相关论文: SS-DPPN: A self-supervised dual-path foundation mo…

200 篇论文

Noise suppression in seismic data processing is a crucial research focus for enhancing subsequent imaging and reservoir prediction. Deep learning has shown promise in computer vision and holds significant potential for seismic data…

地球物理 · 物理学 2024-08-06 Fei Li , Zhenbin Xia , Dawei Liu , Xiaokai Wang , Wenchao Chen , Juan Chen , Leiming Xu

In many signal processing applications, metadata may be advantageously used in conjunction with a high dimensional signal to produce a desired output. In the case of classical Sound Source Localization (SSL) algorithms, information from a…

声音 · 计算机科学 2023-08-09 Eric Grinstein , Vincent W. Neo , Patrick A. Naylor

Generalized Category Discovery (GCD) aims to recognize both known and novel categories from a set of unlabeled data, based on another dataset labeled with only known categories. Without considering differences between known and novel…

计算与语言 · 计算机科学 2023-03-16 Wenbin An , Feng Tian , Qinghua Zheng , Wei Ding , QianYing Wang , Ping Chen

Sleep is essential for maintaining human health and quality of life. Analyzing physiological signals during sleep is critical in assessing sleep quality and diagnosing sleep disorders. However, manual diagnoses by clinicians are…

信号处理 · 电气工程与系统科学 2025-10-02 Cheol-Hui Lee , Hakseung Kim , Byung C. Yoon , Dong-Joo Kim

Over the last few years, deep learning has grown in popularity for speaker verification, identification, and diarization. Inarguably, a significant part of this success is due to the demonstrated effectiveness of their speaker…

声音 · 计算机科学 2022-10-07 Yehoshua Dissen , Felix Kreuk , Joseph Keshet

Self-supervised pretraining (SSP) has been recognized as a method to enhance prediction accuracy in various downstream tasks. However, its efficacy for DNA sequences remains somewhat constrained. This limitation stems primarily from the…

机器学习 · 计算机科学 2024-05-15 Tong Yu , Lei Cheng , Ruslan Khalitov , Erland Brandser Olsson , Zhirong Yang

We propose a novel method to model hierarchical metrical structures for both symbolic music and audio signals in a self-supervised manner with minimal domain knowledge. The model trains and inferences on beat-aligned music signals and…

声音 · 计算机科学 2023-01-26 Junyan Jiang , Gus Xia

Mixed-signal analog/digital circuits emulate spiking neurons and synapses with extremely high energy efficiency, an approach known as "neuromorphic engineering". However, analog circuits are sensitive to process-induced variation among…

机器学习 · 计算机科学 2022-09-13 Julian Büchel , Dmitrii Zendrikov , Sergio Solinas , Giacomo Indiveri , Dylan R. Muir

Several review papers summarize cardiac imaging and DL advances, few works connect this overview to a unified and reproducible experimental benchmark. In this study, we combine a focused review of cardiac ultrasound segmentation literature…

计算机视觉与模式识别 · 计算机科学 2026-01-06 Zahid Ullah , Muhammad Hilal , Eunsoo Lee , Dragan Pamucar , Jihie Kim

Standard fine-tuning of pre-trained audio models couples representation learning with classifier training, which can obscure the true quality of the learned representations. In this work, we advocate for a disentangled two-stage framework…

声音 · 计算机科学 2025-09-23 Yang Wang , Qibin Liang , Chenghao Xiao , Yizhi Li , Noura Al Moubayed , Chenghua Lin

This paper proposes handling training data sparsity in speech-based automatic depression detection (SDD) using foundation models pre-trained with self-supervised learning (SSL). An analysis of SSL representations derived from different…

计算与语言 · 计算机科学 2023-07-07 Wen Wu , Chao Zhang , Philip C. Woodland

This work presents a novel label-efficient selfsupervised representation learning-based approach for classifying diabetic retinopathy (DR) images in cross-domain settings. Most of the existing DR image classification methods are based on…

图像与视频处理 · 电气工程与系统科学 2023-04-25 Ekta Gupta , Varun Gupta , Muskaan Chopra , Prakash Chandra Chhipa , Marcus Liwicki

Recent advances in supervised deep learning methods are enabling remote measurements of photoplethysmography-based physiological signals using facial videos. The performance of these supervised methods, however, are dependent on the…

计算机视觉与模式识别 · 计算机科学 2021-12-15 Hao Wang , Euijoon Ahn , Jinman Kim

A unified self-supervised and supervised deep learning framework for PET image reconstruction is presented, including deep-learned filtered backprojection (DL-FBP) for sinograms, deep-learned backproject then filter (DL-BPF) for…

图像与视频处理 · 电气工程与系统科学 2023-02-28 Andrew J. Reader

This paper introduces GraFPrint, an audio identification framework that leverages the structural learning capabilities of Graph Neural Networks (GNNs) to create robust audio fingerprints. Our method constructs a k-nearest neighbor (k-NN)…

声音 · 计算机科学 2025-01-27 Aditya Bhattacharjee , Shubhr Singh , Emmanouil Benetos

Aligning physiological parameter labels with large-scale photoplethysmographic (PPG) data for deep learning is challenging and resource-intensive. While self-supervised representation learning (SSRL) can handle limited annotated data, the…

信号处理 · 电气工程与系统科学 2026-04-28 Zexing Zhang , Huimin Lu , Songzhe Ma , Jianzhong Peng , Chenglin Lin , Niya Li , Bingwang Dong

Protein phosphorylation provides a dynamic readout of cellular signaling yet remains difficult to detect at low abundance and stoichiometry. Single-molecule surface-enhanced Raman spectroscopy (SM-SERS) using particle-in-pore plasmonic…

介观与纳米尺度物理 · 物理学 2026-04-09 Mulusew W. Yaltaye , Yingqi Zhao , Kuo Zhan , Vahid Farrahi , Jian-An Huang

Several diseases of parkinsonian syndromes present similar symptoms at early stage and no objective widely used diagnostic methods have been approved until now. Positron emission tomography (PET) with $^{18}$F-FDG was shown to be able to…

The need to selectively and efficiently erase learned information from deep neural networks is becoming increasingly important for privacy, regulatory compliance, and adaptive system design. We introduce Graph-Propagated Projection…

计算机视觉与模式识别 · 计算机科学 2026-04-30 Shreyansh Pathak , Jyotishman Das

While supervised learning has achieved remarkable success, obtaining large-scale labeled datasets in biomedical imaging is often impractical due to high costs and the time-consuming annotations required from radiologists. Semi-supervised…

图像与视频处理 · 电气工程与系统科学 2024-01-19 Yuanbin Chen , Tao Wang , Hui Tang , Longxuan Zhao , Ruige Zong , Shun Chen , Tao Tan , Xinlin Zhang , Tong Tong