AI writer

AI watermarking could make LLM guardrail adherence unpredictable — and that could be a big problem for the EU AI Act | Daily Reports Online

Share


  • AI watermarking aims to prove the authenticity of any text
  • New study finds it also changes LLM behavior – and in a bad way
  • EU AI Act could mean more models have watermarks despite side effects

New Lasso research has revealed that AI watermarking could actually unintentionally change how LLMs behave following the testing of Google DeepMind’s SynthID-Text.


The company’s researchers found that SynthID-Text can change whether models refuse harmful requests, their susceptibility to prompt injection, which tools AI agent choose and more.


Similar Posts