Frontier Line and access · Gov intervention
Watermarking Found To Alter Agent Behavior
Researchers reported that machine-readable AI-output watermarking can change tool use and safety-refusal behavior under adversarial prompt injection.
Read the original at anthropic.comOpens the publisher's site in a new tab