中文
相关论文

相关论文: Intervention-Based Self-Supervised Learning: A Cau…

200 篇论文

Imaging Photoplethysmography (iPPG), an optical procedure which recovers a human's blood volume pulse (BVP) waveform using pixel readout from a camera, is an exciting research field with many researchers performing clinical studies of iPPG…

计算机视觉与模式识别 · 计算机科学 2026-04-14 Vineet R. Shenoy , Cheng Peng , Rama Chellappa , Yu Sun

The task of medical image recognition is notably complicated by the presence of varied and multiple pathological indications, presenting a unique challenge in multi-label classification with unseen labels. This complexity underlines the…

计算机视觉与模式识别 · 计算机科学 2024-09-16 Yaoqin Ye , Junjie Zhang , Hongwei Shi

Current foundation model for photoplethysmography (PPG) signals is challenged by the intrinsic redundancy and noise of the signal. Standard masked modeling often yields trivial solutions while contrastive methods lack morphological…

机器学习 · 计算机科学 2026-01-30 Zongheng Guo , Tao Chen , Yang Jiao , Yi Pan , Xiao Hu , Manuela Ferrario

Remote photoplethysmography (rPPG) enables non-contact measurement of cardiac pulse signals by analyzing subtle color changes in facial videos. Nevertheless, extracting rPPG signals remains challenging because of their extremely weak signal…

图像与视频处理 · 电气工程与系统科学 2026-05-22 Kosuke Kurihara , Yoshihiro Maeda , Daisuke Sugimura , Takayuki Hamamoto

Deep learning highly relies on the amount of annotated data. However, annotating medical images is extremely laborious and expensive. To this end, self-supervised learning (SSL), as a potential solution for deficient annotated data,…

计算机视觉与模式识别 · 计算机科学 2020-06-11 Jiuwen Zhu , Yuexiang Li , Yifan Hu , S. Kevin Zhou

Single-pixel imaging (SPI) offers a cost-effective route to hyperspectral acquisition but struggles to recover high-fidelity spatial and spectral details under extremely low sampling rates, a severely ill-posed inverse problem. While deep…

计算机视觉与模式识别 · 计算机科学 2026-04-07 Hao Zhang , Bilige Xu , Lichen Wei , Xu Ma , Wenyi Ren

In this paper, we propose a method that learns a general representation of periodic signals from unlabeled facial videos by capturing subtle changes in skin tone over time. The proposed framework employs the video masked autoencoder to…

计算机视觉与模式识别 · 计算机科学 2025-06-30 Jiho Choi , Sang Jun Lee

Remote photoplethysmography (rPPG) enables non-contact, continuous monitoring of physiological signals and offers a practical alternative to traditional health sensing methods. Although rPPG is promising for daily health monitoring, its…

计算机视觉与模式识别 · 计算机科学 2025-11-04 Xulin Ma , Jiankai Tang , Zhang Jiang , Songqin Cheng , Yuanchun Shi , Dong LI , Xin Liu , Daniel McDuff , Xiaojing Liu , Yuntao Wang

The COVID-19 pandemic has underscored the need for low-cost, scalable approaches to measuring contactless vital signs, either during initial triage at a healthcare facility or virtual telemedicine visits. Remote photoplethysmography (rPPG)…

图像与视频处理 · 电气工程与系统科学 2024-11-26 Sandeep Nagar , Mark Hasegawa-Johnson , David G. Beiser , Narendra Ahuja

Remote photoplethysmography (rPPG) is a method for measuring a subjects heart rate remotely using a camera. Factors such as subject movement, ambient light level, makeup etc. complicate such measurements by distorting the observed pulse.…

计算机视觉与模式识别 · 计算机科学 2025-02-05 Alexey Protopopov

Multi-view 3D reconstruction methods remain highly sensitive to photometric inconsistencies arising from camera optical characteristics and variations in image signal processing (ISP). Existing mitigation strategies such as per-frame latent…

计算机视觉与模式识别 · 计算机科学 2026-04-08 Isaac Deutsch , Nicolas Moënne-Loccoz , Gavriel State , Zan Gojcic

Prompt learning is an effective method to customize Vision-Language Models (VLMs) for various downstream tasks, involving tuning very few parameters of input prompt tokens. Recently, prompt pretraining in large-scale dataset (e.g.,…

计算机视觉与模式识别 · 计算机科学 2024-09-11 Zhenyuan Chen , Lingfeng Yang , Shuo Chen , Zhaowei Chen , Jiajun Liang , Xiang Li

Managing the dynamic regions in the photometric loss formulation has been a main issue for handling the self-supervised depth estimation problem. Most previous methods have alleviated this issue by removing the dynamic regions in the…

计算机视觉与模式识别 · 计算机科学 2022-05-23 Geonho Cha , Ho-Deok Jang , Dongyoon Wee

Contrastive learning with the nearest neighbor has proved to be one of the most efficient self-supervised learning (SSL) techniques by utilizing the similarity of multiple instances within the same class. However, its efficacy is…

计算机视觉与模式识别 · 计算机科学 2025-04-25 Dewen Zeng , Yawen Wu , Xinrong Hu , Xiaowei Xu , Yiyu Shi

Snapshot compressive imaging (SCI) captures high-dimensional data efficiently by compressing it into two-dimensional observations and reconstructing high-dimensional data from two-dimensional observations with various algorithms. The…

图像与视频处理 · 电气工程与系统科学 2025-03-06 Takashi Matsuda , Ryo Hayakawa , Youji Iiguni

Remote photoplethysmography (rPPG), which aims at measuring heart activities and physiological signals from facial video without any contact, has great potential in many applications (e.g., remote healthcare and affective computing). Recent…

计算机视觉与模式识别 · 计算机科学 2022-05-24 Zitong Yu , Yuming Shen , Jingang Shi , Hengshuang Zhao , Philip Torr , Guoying Zhao

Causal probing methods aim to test and control how internal representations influence the behavior of generative models. In causal probing, an intervention modifies hidden states so that a property takes on a different value. Most existing…

人工智能 · 计算机科学 2026-05-11 Sadegh Khorasani , Saber Salehkaleybar , Negar Kiyavash , Matthias Grossglauser

This paper presents an approach to assess the perfusion of visible human tissue from RGB video files. We propose metrics derived from remote photoplethysmography (rPPG) signals to detect whether a tissue is adequately supplied with blood.…

计算机视觉与模式识别 · 计算机科学 2022-12-26 Benjamin Kossack , Eric Wisotzky , Peter Eisert , Sebastian P. Schraven , Brigitta Globke , Anna Hilsmann

Existing Image-Text Sentiment Analysis (ITSA) methods may suffer from inconsistent intra-modal and inter-modal sentiment relationships. Therefore, we develop a method that balances before fusing to solve the issue of vision-language…

多媒体 · 计算机科学 2026-02-25 Jiesheng Wu , Shengrong Li

Accurate depth estimation enhances endoscopy navigation and diagnostics, but obtaining ground-truth depth in clinical settings is challenging. Synthetic datasets are often used for training, yet the domain gap limits generalization to real…

计算机视觉与模式识别 · 计算机科学 2025-04-25 Xinqi Xiong , Andrea Dunn Beltran , Jun Myeong Choi , Marc Niethammer , Roni Sengupta