Tag
This article analyzes how LLM watermarking, specifically SynthID-Text, impacts AI agent behavior by causing sampling drift that affects model refusals and tool calling, with implications for AI safety and regulatory compliance.
Research finds that AI text watermarking alters language model responses to harmful prompts, potentially increasing vulnerability to adversarial attacks and affecting agent behavior.