Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

FlowMark: Mask-Guided Video Watermarking

About

We present FlowMark, a video watermarking framework guided by automatically predicted object masks. In contrast to prior region-based approaches that require user-supplied mask guidance, FlowMark learns to identify optimal regions for watermark embedding through a dedicated Mask Predictor network. Our end-to-end trainable architecture combines region-aware encoding with noise-augmented training to ensure robustness against compression, geometric transformations, and content variation, while preserving high perceptual quality. Our content-adaptive masking keeps watermark signals coherent with natural video dynamics, effectively eliminating perceptual flicker. Beyond compression robustness, FlowMark maintains reliable watermark recovery under video-native temporal edits (e.g., frame swap, insertion, deletion, resampling, and interpolation) and real-world social media distribution pipelines (e.g., YouTube and Facebook re-encoding). Experimental results on both image and video datasets show that FlowMark reliably embeds $128$-bit messages with up to $50.08$ dB PSNR, offering strong performance for content provenance, temporal authenticity verification, and video integrity protection.

Vishal Asnani, Shruti Agarwal, John Collomosse• 2026

Related benchmarks

TaskDatasetResultRank
Watermarking RobustnessSA-V video (test)
Value Bit Accuracy98
8
Watermarking RobustnessSA-1b image (test)
Value Bit Accuracy99
8
WatermarkingSA-1b 1.0
Bit Accuracy100
8
WatermarkingSA-V 1.0 (test)
Bit Accuracy100
8
Video Message EmbeddingVideo Frames 256 x 256 VideoSeal setup
Embedding GFLOPs12.7
6
Video Message Extraction256 x 256 Video Frames VideoSeal setup
GFLOPs (Extraction)17.6
6
Showing 6 of 6 rows

Other info

Follow for update