English
Related papers

Related papers: Training Robust Deep Physiological Measurement Mod…

200 papers

Recent work has shown that a person's sympathetic arousal can be estimated from facial videos alone using basic signal processing. This opens up new possibilities in the field of telehealth and stress management, providing a non-invasive…

Computer Vision and Pattern Recognition · Computer Science 2024-10-29 Björn Braun , Daniel McDuff , Tadas Baltrusaitis , Paul Streli , Max Moebus , Christian Holz

Personalized computed tomography (CT) dosimetry has great potential in assessing patient-specific radiation exposure, supporting risk assessment, and optimizing clinical protocols. The aim of this study is to evaluate the potential of…

Medical Physics · Physics 2026-01-15 Marie-Luise Kuhlmann , Jörg Martin , Stefan Pojtinger

Falls represent a significant cause of injury among the elderly population. Extensive research has been devoted to the utilization of wearable IMU sensors in conjunction with machine learning techniques for fall detection. To address the…

Quantitative Methods · Quantitative Biology 2023-10-18 Jie Tang , Bin He , Junkai Xu , Tian Tan , Zhipeng Wang , Yanmin Zhou , Shuo Jiang

The proliferation of generative video technologies has intensified the need for reliable methods to detect and characterize synthetic media. To address this challenge, we organized the \href{https://safe-video-2025.dsri.org}{SAFE: Synthetic…

Large fingerprint datasets, while important for training and evaluation, are time-consuming and expensive to collect and require strict privacy measures. Researchers are exploring the use of synthetic fingerprint data to address these…

Computer Vision and Pattern Recognition · Computer Science 2026-01-15 Syed Konain Abbas , Sandip Purnapatra , M. G. Sarwar Murshed , Conor Miller-Lynch , Lambert Igene , Soumyabrata Dey , Stephanie Schuckers , Faraz Hussain

Real-world talking faces often accompany with natural head movement. However, most existing talking face video generation methods only consider facial animation with fixed head pose. In this paper, we address this problem by proposing a…

Computer Vision and Pattern Recognition · Computer Science 2020-03-06 Ran Yi , Zipeng Ye , Juyong Zhang , Hujun Bao , Yong-Jin Liu

To date, research on sensor-equipped mobile devices has primarily focused on the purely supervised task of human activity recognition (walking, running, etc), demonstrating limited success in inferring high-level health outcomes from…

Machine Learning · Computer Science 2020-11-10 Dimitris Spathis , Ignacio Perez-Pozuelo , Soren Brage , Nicholas J. Wareham , Cecilia Mascolo

In this paper, we propose to pre-train audio encoders using synthetic patterns instead of real audio data. Our proposed framework consists of two key elements. The first one is Masked Autoencoder (MAE), a self-supervised learning framework…

Audio and Speech Processing · Electrical Eng. & Systems 2024-10-02 Yuchi Ishikawa , Tatsuya Komatsu , Yoshimitsu Aoki

Deep learning-based methods for video pedestrian detection and tracking require large volumes of training data to achieve good performance. However, data acquisition in crowded public environments raises data privacy concerns -- we are not…

Computer Vision and Pattern Recognition · Computer Science 2021-08-24 Matteo Fabbri , Guillem Braso , Gianluca Maugeri , Orcun Cetintas , Riccardo Gasparini , Aljosa Osep , Simone Calderara , Laura Leal-Taixe , Rita Cucchiara

Video matting has traditionally been limited by the lack of high-quality ground-truth data. Most existing video matting datasets provide only human-annotated imperfect alpha and foreground annotations, which must be composited to background…

Computer Vision and Pattern Recognition · Computer Science 2025-08-12 Yongtao Ge , Kangyang Xie , Guangkai Xu , Mingyu Liu , Li Ke , Longtao Huang , Hui Xue , Hao Chen , Chunhua Shen

Foundation models such as Segment Anything Model 2 (SAM 2) exhibit strong generalization on natural images and videos but perform poorly on medical data due to differences in appearance statistics, imaging physics, and three-dimensional…

Image and Video Processing · Electrical Eng. & Systems 2026-01-21 Satrajit Chakrabarty , Sourya Sengupta , Gopal Avinash , Ravi Soni

We present an overview and evaluation of a new, systematic approach for generation of highly realistic, annotated synthetic data for training of deep neural networks in computer vision tasks. The main contribution is a procedural world…

Computer Vision and Pattern Recognition · Computer Science 2017-10-19 Apostolia Tsirikoglou , Joel Kronander , Magnus Wrenninge , Jonas Unger

This study proposes a method for simulating signals received by frequency-modulated continuous-wave radar during respiratory monitoring, using human body geometry and displacement data acquired via a depth camera. Unlike previous studies…

Signal Processing · Electrical Eng. & Systems 2025-07-21 Kimitaka Sumi , Takuya Sakamoto

This report introduces VitalLens 2.0, a new deep learning model for estimating physiological signals from face video. This new model demonstrates a significant leap in accuracy for remote photoplethysmography (rPPG), enabling the robust…

Computer Vision and Pattern Recognition · Computer Science 2025-11-03 Philipp V. Rouast

In this paper, we propose a novel face synthesis approach that can generate an arbitrarily large number of synthetic images of both real and synthetic identities. Thus a face image dataset can be expanded in terms of the number of…

Computer Vision and Pattern Recognition · Computer Science 2017-04-26 Sandipan Banerjee , John S. Bernhard , Walter J. Scheirer , Kevin W. Bowyer , Patrick J. Flynn

Despite considerable efforts to enhance the generalization of 3D pose estimators without costly 3D annotations, existing data augmentation methods struggle in real world scenarios with diverse human appearances and complex poses. We propose…

Computer Vision and Pattern Recognition · Computer Science 2025-03-18 ChangHee Yang , Hyeonseop Song , Seokhun Choi , Seungwoo Lee , Jaechul Kim , Hoseok Do

Advances in machine learning have enabled the creation of realistic synthetic videos known as deepfakes. As deepfakes proliferate, concerns about rapid spread of disinformation and manipulation of public perception are mounting. Despite the…

Human-Computer Interaction · Computer Science 2026-03-31 David Wegmann , Emil Stevnsborg , Søren Knudsen , Luca Rossi , Aske Mottelson

Facial expression datasets remain limited in scale due to the subjectivity of annotations and the labor-intensive nature of data collection. This limitation poses a significant challenge for developing modern deep learning-based facial…

Computer Vision and Pattern Recognition · Computer Science 2025-08-13 Xilin He , Cheng Luo , Xiaole Xian , Bing Li , Muhammad Haris Khan , Zongyuan Ge , Weicheng Xie , Siyang Song , Linlin Shen , Bernard Ghanem , Xiangyu Yue

Multi-animal pose estimation is essential for studying animals' social behaviors in neuroscience and neuroethology. Advanced approaches have been proposed to support multi-animal estimation and achieve state-of-the-art performance. However,…

Computer Vision and Pattern Recognition · Computer Science 2022-04-15 Ari Blau , Christoph Gebhardt , Andres Bendesky , Liam Paninski , Anqi Wu

High-quality, large-scale data is essential for robust deep learning models in medical applications, particularly ultrasound image analysis. Diffusion models facilitate high-fidelity medical image generation, reducing the costs associated…

Image and Video Processing · Electrical Eng. & Systems 2024-04-01 Pooria Ashrafian , Milad Yazdani , Moein Heidari , Dena Shahriari , Ilker Hacihaliloglu
‹ Prev 1 8 9 10 Next ›