A Fine-Grained Semantic Steganography Framework with Cross-Modal Drift Regularization
Steganography using deep learning can preserve pixel-level image quality while still changing object, attribute, or relational information in captions generated by vision-language models (VLMs). This caption drift creates a detection channel that is not measured by global image-embedding similarity alone. This paper pr...