Generative Text Steganography with Large Language Model
Jiaxuan Wu, Zhengxian Wu, Yiming Xue, Juan Wen, Wanli Peng
摘要
Recent advances in large language models (LLMs) have blurred the boundary of high-quality text generation between humans and machines, which is favorable for generative text steganography. Currently, advanced steganographic mapping is not suitable for LLMs since most users are restricted to accessing only the black-box API or user interface of the LLMs, thereby lacking access to the training vocabulary and its sampling probabilities. In this paper, we explore a black-box generative text steganographic method based on the user interfaces of large language models, which is called LLM-Stega. The main goal of LLM-Stega is to ensure secure covert communication between Alice (sender) and Bob (receiver) by using the user interfaces of LLMs. Specifically, We first construct a keyword set and design a new encrypted steganographic mapping to embed secret messages. Furthermore, an optimization mechanism based on reject sampling is proposed to guarantee accurate extraction of secret messages and rich semantics of generated stego texts. Comprehensive experiments demonstrate that the proposed LLM-Stega outperforms current state-of-the-art methods.
• Security and privacy → Human and societal aspects of security and privacy; • Computing methodologies → Natural language generation.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- All That Glitters Is Not Gold: Key-Secured 3D Secrets within 3D Gaussian SplattingYan Ren, Shilin Lu, Adams Wai-Kin KongICLR 2026 · 被引用 16 次
- All Code, No Thought: Language Models Struggle to Reason in Ciphered LanguageShiyuan Guo, Henry Sleight, Fabien RogerICLR 2026 · 被引用 5 次
- STEAD: Robust Provably Secure Linguistic Steganography with Diffusion Language ModelYuang Qi, Na Zhao, Qiyi Yao, Benlong Wu 等NeurIPS 2025 · 被引用 4 次
- LLMs Can Hide Text in Other Text of the Same LengthAntonio Norelli, Michael M. BronsteinICLR 2026 · 被引用 3 次
- ImF: Embedding an Implicit Fingerprint in Your Large Language ModelsJiaxuan Wu, Wanli Peng, Hang Fu, Yiming Xue 等ACL 2026
它引用的顶会 Paper10
- Symmetric Cross Entropy for Robust Learning With Noisy LabelsYisen Wang, Xingjun Ma, Zaiyi Chen, Yuan Luo 等ICCV 2019 · 被引用 1,125 次
- Generating Training Data with Language Models: Towards Zero-Shot Language UnderstandingYu Meng, Jiaxin Huang, Yu Zhang, Jiawei HanNeurIPS 2022 · 被引用 309 次
- CRoSS: Diffusion Model Makes Controllable, Robust and Secure Image SteganographyJiwen Yu, Xuanyu Zhang, Youmin Xu, Jian ZhangNeurIPS 2023 · 被引用 146 次
- ZeroGen: Efficient Zero-shot Learning via Dataset GenerationJiacheng Ye, Jiahui Gao, Qintong Li, Hang Xu 等EMNLP 2022 · 被引用 96 次
- Language Models are Realistic Tabular Data GeneratorsVadim Borisov, Kathrin Seßler, Tobias Leemann, Martin Pawelczyk 等ICLR 2023 · 被引用 45 次
相关 Paper
- Provable Secure Steganography Based on Adaptive Dynamic SamplingKaiyi Pang, Minhao BaiUSENIX Security 2026 · 被引用 4 次
- TrojanStego: Your Language Model Can Secretly Be A Steganographic Privacy Leaking AgentDominik Meier, Jan Philip Wahle, Paul Röttger, Terry Ruas 等EMNLP 2025
- TrojLLM: A Black-box Trojan Prompt Attack on Large Language ModelsJiaqi Xue, Mengxin Zheng, Ting Hua, Yilin Shen 等NeurIPS 2023 · 被引用 63 次
- Early Signs of Steganographic Capabilities in Frontier LLMsArtur Zolkowski, Kei Nishimura-Gasparian, Robert McCarthy, Roland S. Zimmermann 等ICLR 2026 · 被引用 22 次
- A Content-Preserving Secure Linguistic SteganographyLingyun Xiang, Chengfu Ou, Xu He, Zhongliang Yang 等AAAI 2026
