SPAct: Self-supervised Privacy Preservation for Action Recognition

About

Visual private information leakage is an emerging key issue for the fast growing applications of video understanding like activity recognition. Existing approaches for mitigating privacy leakage in action recognition require privacy labels along with the action labels from the video dataset. However, annotating frames of video dataset for privacy labels is not feasible. Recent developments of self-supervised learning (SSL) have unleashed the untapped potential of the unlabeled data. For the first time, we present a novel training framework which removes privacy information from input video in a self-supervised manner without requiring privacy labels. Our training framework consists of three main components: anonymization function, self-supervised privacy removal branch, and action recognition branch. We train our framework using a minimax optimization strategy to minimize the action recognition cost function and maximize the privacy cost function through a contrastive self-supervised loss. Employing existing protocols of known-action and privacy attributes, our framework achieves a competitive action-privacy trade-off to the existing state-of-the-art supervised methods. In addition, we introduce a new protocol to evaluate the generalization of learned the anonymization function to novel-action and privacy attributes and show that our self-supervised framework outperforms existing supervised methods. Code available at: https://github.com/DAVEISHAN/SPAct

Ishan Rajendrakumar Dave, Chen Chen, Mubarak Shah• 2022

Related benchmarks

Task	Dataset	Result
Action Recognition	Kinetics-400	Top-1 Acc46.93	505
Action Recognition	UCF101	--	433
Visual Question Answering	OK-VQA	Accuracy56.5	331
Visual Question Answering	OK-VQA (test)	Accuracy55.9	327
Video Action Recognition	UCF101 (test)	Top-1 Acc62.03	46
Emotion Recognition	CREMA-D (test)	Accuracy74.59	37
Emotion Recognition	RAVDESS (test)	Accuracy0.7731	29
Action Recognition	HMDB51 VISPR	Top-1 Accuracy51.7	24
Action Recognition	UCF101 VISPR	Top-1 Accuracy63.1	24
Action Recognition	PA-HMDB	Top-1 Accuracy43.1	19

Showing 10 of 37 rows

Other info

Code

Follow for update

@wizwand_team Discord