anthropic Anthropic News ·

Anthropic implements text watermarking in Claude to comply with EU AI Act

aigaengineer
feature announcement

Anthropic has implemented text watermarking in future Claude models to help determine the likelihood of AI involvement in generated content. This change, effective as of August 2, is a response to the EU AI Act, which mandates AI providers operating in its market to mark AI-generated text. The watermarking method is based on Google DeepMind's SynthID-Text, discreetly altering word choices without impacting output quality, cost, or readability. It does not add hidden characters or identifying information, and is indistinguishable to human readers.

  • Introduction of text watermarking for Claude models
  • How Claude's text watermarking works
  • No impact on Claude's output quality or cost
  • Watermarking method and its limitations
  • Watermarking on edited text and code
Notes (5)
  • Introduction of text watermarking for Claude models

    Future Claude models will generate watermarked text to indicate AI involvement, a measure implemented to comply with the EU AI Act's August 2 mandate for AI providers. This initiative is being adopted by Anthropic and several other major AI providers.

  • How Claude's text watermarking works

    The watermarking method subtly influences low-stakes word choices made by the large language model, creating an undetectable pattern that can be identified with a key. This process does not alter the meaning or readability of the generated text, ensuring it's indistinguishable to readers.

  • No impact on Claude's output quality or cost

    Watermarking does not affect the content, creativity, or readability of Claude's outputs. It does not add hidden characters, require extra tokens, or increase generation costs, ensuring no practical impact on quality or expense.

  • Watermarking method and its limitations

    Claude uses a version of Google DeepMind's SynthID-Text approach. While effective for assessing the likelihood of Claude's involvement, it cannot confirm human authorship, identify other AIs, and is less effective on short, highly factual, or lightly edited passages due to fewer word choices.

  • Watermarking on edited text and code

    The watermark applies only to words Claude chooses. For lightly edited human text or code requiring exact outputs (e.g., factual statements, functional code), there's little to no impact, as the model makes fewer or no low-stakes choices for watermarking.

Read the original announcement →

https://www.anthropic.com/news/claude-text-watermark

Related releases