The Provenance Tax: Understanding the Impact of LLM Watermarking on AI Agent Behavior
Summary
The article analyzes how LLM watermarking (SynthID-Text) used for provenance can alter model and agent behavior, including tool calls and refusals under prompt injection. It presents empirical findings on sampling drift, tool-call churn, and safety implications, and argues for thorough evaluation of watermark configurations in deployed agents.