AI watermarks may change how agents use tools and handle malicious prompts
Techstrong.ai reports that Lasso Security research found AI watermarking can alter tool calls, refusal behavior and security outcomes in unexpected ways.
TLDR
Techstrong.ai reports that Lasso Security found AI watermarks can affect tool calls, arguments and how models respond to malicious prompts in agentic systems. The outlet also describes potential changes to refusal behavior and security outcomes.
AI watermarks may change how agents use tools and handle malicious prompts
Techstrong.ai reports that Lasso Security research found AI watermarking can alter tool calls, refusal behavior and security outcomes in unexpected ways.
TLDR
Techstrong.ai reports that Lasso Security found AI watermarks can affect tool calls, arguments and how models respond to malicious prompts in agentic systems. The outlet also describes potential changes to refusal behavior and security outcomes.