Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

Addressing Segmentation Ambiguity in Neural Linguistic Steganography

About

Previous studies on neural linguistic steganography, except Ueoka et al. (2021), overlook the fact that the sender must detokenize cover texts to avoid arousing the eavesdropper's suspicion. In this paper, we demonstrate that segmentation ambiguity indeed causes occasional decoding failures at the receiver's side. With the near-ubiquity of subwords, this problem now affects any language. We propose simple tricks to overcome this problem, which are even applicable to languages without explicit word boundaries.

Jumon Nozaki, Yugo Murawaki• 2022

Related benchmarks

TaskDatasetResultRank
Linguistic SteganographyChinese IMDB
Avg PPL (bits/token)21.1
21
Tokenization DisambiguationIMDb English
Average Perplexity (PPL)19.93
21
Showing 2 of 2 rows

Other info

Follow for update