AI model watermarking changes agent behavior
Lasso Security says watermarking used for AI provenance can change AI agents’ tool-calling accuracy and refusal behavior, and under adversarial prompt injection it increased attack success rates by making models less likely to refuse harmful requests.
Sep 17, 2026 ·
The Register










