Logo image
Multi-task Learning for Simultaneous Video Generation and Remote Photoplethysmography Estimation
Conference paper   Peer reviewed

Multi-task Learning for Simultaneous Video Generation and Remote Photoplethysmography Estimation

Yun-Yun Tsou, Yi-An Lee and Chiou-Ting Hsu
Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics), Vol.12626 LNCS, pp.392-407
2021

Abstract

Theoretical Computer Science Computer Science (all)
Remote photoplethysmography (rPPG) is a contactless method for estimating physiological signals from facial videos. Without large supervised datasets, learning a robust rPPG estimation model is extremely challenging. Instead of merely focusing on model learning, we believe data augmentation may be of greater importance for this task. In this paper, we propose a novel multi-task learning framework to simultaneously augment training data while learning the rPPG estimation model. We design three networks: rPPG estimation network, Image-to-Video network, and Video-to-Video network, to estimate rPPG signals from face videos, to generate synthetic videos from a source image and a specified rPPG signal, and to generate synthetic videos from a source video and a specified rPPG signal, respectively. Experimental results on three benchmark datasets, COHFACE, UBFC, and PURE, show that our method successfully generates photo-realistic videos and significantly outperforms existing methods with a large margin. (The code is publicly available at https://github.com/YiAnLee/Multi-Task-Learning-for-Simultaneous-VideoGeneration-and-Remote-Photoplethysmography-Estimation ).

Metrics

1 Record Views

Details

Logo image