Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

Anti-Prompt: Image Protection against Text-Guided Image-to-Video Generation

About

Recent advances in Image-to-Video generation allow a single image to be animated into a convincing video under text guidance, raising serious copyright and privacy risks. We propose Anti-Prompt, an image protection approach that injects imperceptible perturbations into an image, inducing visible inconsistencies and structural failures in text-guided I2V generation. Our method is motivated by a simple empirical observation. When text guidance is removed from modern I2V models, generation quality degrades markedly, not only in motion realism but also in subject preservation, structural coherence, and temporal consistency. Building on this insight, Anti-Prompt exploits the model reliance on textual guidance by attenuating text-conditioned interactions during denoising while strengthening visual-only pathways. To further systematically evaluate protection effectiveness, we introduce a Video-LLM-assisted evaluation protocol that provides interpretable, frame-grounded analyses of generation artifacts and inconsistencies. Experiments on two representative I2V architectures demonstrate that our method achieves strong protection performance while improving efficiency and cross-model transferability.

Yeonghwan Song, Chanhui Lee, Jinsoo Park, Jeany Son• 2026

Related benchmarks

TaskDatasetResultRank
Image-to-Video GenerationVBench
Motion Smoothness0.9883
46
Image-to-Video GenerationVBench (test)
Aesthetic Quality Score58.26
34
Robustness to purification defensesCogVideoX (test)
Subject Preservation4.75
6
Robustness to purification defensesLTX-Video (test)
Subject Preservation4.31
6
Image-to-Video Protection EvaluationI2V protection evaluation protocol unseen prompts 1.0
Subject Preservation Score4.26
6
Image-to-Video Protection EvaluationI2V protection evaluation protocol 1.0 (seen prompts)
Subject Preservation4.37
6
Image-to-Video GenerationVBench White-box quantitative results
Subject Fidelity93.64
6
Video Generation EvaluationVBench unseen prompt
Subject Fidelity91.14
6
Computational Efficiency AnalysisI2V Generative Models
Video Generation PFLOPs0.00e+0
4
Imperceptibility EvaluationProtection Image-to-Video (Evaluation Set)
DISTS0.111
4
Showing 10 of 14 rows

Other info

Follow for update