Now You (Still) See Me: Detecting Evasive Steganographic Payloads in LLMs
DGX agentarXiv:2606.09411v1 Announce Type: cross Abstract: Large language models can be fine-tuned to encode prompt-borne secrets into fluent, seemingly benign outputs. This creates a steganographic exfiltrati