中文
相关论文

相关论文: Joint Spatial-Temporal Modeling and Contrastive Le…

200 篇论文

In this work, we present a novel learning-based framework that combines the local accuracy of contrastive learning with the global consistency of geometric approaches, for robust non-rigid matching. We first observe that while contrastive…

计算机视觉与模式识别 · 计算机科学 2022-09-19 Lei Li , Souhaib Attaiki , Maks Ovsjanikov

Vital signs, such as heart rate (HR), heart rate variability (HRV), respiratory rate (RR), are important indicators for a person's health. Vital signs are traditionally measured with contact sensors, and may be inconvenient and cause…

图像与视频处理 · 电气工程与系统科学 2019-11-04 Mingliang Chen , Qiang Zhu , Harrison Zhang , Min Wu , Quanzeng Wang

Foundation models have recently gained attention within the field of machine learning thanks to its efficiency in broad data processing. While researchers had attempted to extend this success to time series models, the main challenge is…

机器学习 · 计算机科学 2023-11-22 Trang H. Tran , Lam M. Nguyen , Kyongmin Yeo , Nam Nguyen , Roman Vaculin

Electronic Health Record (EHR) data has been of tremendous utility in Artificial Intelligence (AI) for healthcare such as predicting future clinical events. These tasks, however, often come with many challenges when using classical machine…

机器学习 · 计算机科学 2021-04-08 Tingyi Wanyan , Jing Zhang , Ying Ding , Ariful Azad , Zhangyang Wang , Benjamin S Glicksberg

Self-supervised learning has proven to be an effective way to learn representations in domains where annotated labels are scarce, such as medical imaging. A widely adopted framework for this purpose is contrastive learning and it has been…

计算机视觉与模式识别 · 计算机科学 2024-02-23 Hugo Figueiras , Helena Aidos , Nuno Cruz Garcia

We propose a supervised contrastive learning framework for video representation learning that leverages temporally global context. We introduce a video to image aggregation strategy that spatially arranges multiple frames from each video…

计算机视觉与模式识别 · 计算机科学 2025-12-16 Shaif Chowdhury , Mushfika Rahman , Greg Hamerly

In this paper, we propose a novel contrastive learning based deep learning framework for patient similarity search using physiological signals. We use a contrastive learning based approach to learn similar embeddings of patients with…

信号处理 · 电气工程与系统科学 2023-08-07 Subangkar Karmaker Shanto , Shoumik Saha , Atif Hasan Rahman , Mohammad Mehedy Masud , Mohammed Eunus Ali

We propose an end-to-end framework to measure people's vital signs including Heart Rate (HR), Heart Rate Variability (HRV), Oxygen Saturation (SpO2) and Blood Pressure (BP) based on the rPPG methodology from the video of a user's face…

计算机视觉与模式识别 · 计算机科学 2022-12-23 Donghao Qiao , Amtul Haq Ayesha , Farhana Zulkernine , Raihan Masroor , Nauman Jaffar

Existing self-supervised learning (SSL) methods primarily learn object-invariant representations but often neglect the spatial structure and relationships among object parts. To address this limitation, we introduce Spatial Prediction (SP),…

计算机视觉与模式识别 · 计算机科学 2026-05-12 Yang Shen , Yusen Cai , Weronika Hryniewska-Guzik , Qing Lin , Mengmi Zhang

Geo-tagged images are publicly available in large quantities, whereas labels such as object classes are rather scarce and expensive to collect. Meanwhile, contrastive learning has achieved tremendous success in various natural image and…

计算机视觉与模式识别 · 计算机科学 2023-05-10 Gengchen Mai , Ni Lao , Yutong He , Jiaming Song , Stefano Ermon

The integration of GNSS data into portable devices has led to the generation of vast amounts of trajectory data, which is crucial for applications such as map-matching. To tackle the limitations of rule-based methods, recent works in deep…

数据库 · 计算机科学 2026-03-26 Anjun Gao , Zhenglin Wan , Pingfu Chao , Shunyu Yao

We propose a novel heart rate (HR) estimation method from facial videos that dynamically adapts the HR pulse extraction algorithm to separately deal with noise from 'rigid' head motion and 'non-rigid' facial expression. We first identify…

信号处理 · 电气工程与系统科学 2019-05-30 Utkarsh Sharma , Terumi Umematsu , Masanori Tsujikawa , Yoshifumi Onishi

We present a method for finding cross-modal space-time correspondences. Given two images from different visual modalities, such as an RGB image and a depth map, our model identifies which pairs of pixels correspond to the same physical…

计算机视觉与模式识别 · 计算机科学 2025-06-04 Ayush Shrivastava , Andrew Owens

This study evaluates remote Photopletismography (rPPG) algorithms, Spatial Subspace Rotation (2SR), Chrominance-based method (CHROM), Plane-Orthogonal-to-Skin (POS), and Principal Component Analysis (PCA), applied to selected…

图像与视频处理 · 电气工程与系统科学 2026-05-26 Đorđe D. Nešković , Nadica Miljković

We consider the problem of predicting how the likelihood of an outcome of interest for a patient changes over time as we observe more of the patient data. To solve this problem, we propose a supervised contrastive learning framework that…

机器学习 · 计算机科学 2024-04-16 Shahriar Noroozizadeh , Jeremy C. Weiss , George H. Chen

Subtle periodic signals such as blood volume pulse and respiration can be extracted from RGB video, enabling remote health monitoring at low cost. Advancements in remote pulse estimation -- or remote photoplethysmography (rPPG) -- are…

计算机视觉与模式识别 · 计算机科学 2023-03-15 Jeremy Speth , Nathan Vance , Patrick Flynn , Adam Czajka

Finding point-level correspondences is a fundamental problem in ultrasound (US), since it can enable US landmark tracking for intraoperative image guidance in different surgeries, including head and neck. Most existing US tracking methods,…

计算机视觉与模式识别 · 计算机科学 2024-08-01 Wanwen Chen , Adam Schmidt , Eitan Prisman , Septimiu E Salcudean

Joint Detection and Embedding (JDE) trackers have demonstrated excellent performance in Multi-Object Tracking (MOT) tasks by incorporating the extraction of appearance features as auxiliary tasks through embedding Re-Identification task…

计算机视觉与模式识别 · 计算机科学 2024-08-07 Yunfei Zhang , Chao Liang , Jin Gao , Zhipeng Zhang , Weiming Hu , Stephen Maybank , Xue Zhou , Liang Li

Self-supervised learning (SSL) has emerged as a powerful approach to learning representations, particularly in the field of computer vision. However, its application to dependent data, such as temporal and spatio-temporal domains, remains…

机器学习 · 计算机科学 2025-10-01 Alexander Marusov , Aleksandr Yugay , Alexey Zaytsev

Ambiguities in data and problem constraints can lead to diverse, equally plausible outcomes for a machine learning task. In beat and downbeat tracking, for instance, different listeners may adopt various rhythmic interpretations, none of…

声音 · 计算机科学 2025-10-30 Antonin Gagnere , Slim Essid , Geoffroy Peeters