← Back to Blogs
Scott Aaronson

Anthropic’s LLM watermarking

Here is a 3-paragraph summary of the blog post for mathopen.com:

Anthropic has announced that it is now watermarking the outputs of its Claude AI models. The watermarking scheme they are using is based on Google's SynthID system, which itself traces back to an earlier foundational proposal called the Gumbel Softmax watermarking scheme.

That original Gumbel Softmax scheme was first proposed by the blog's author while working at OpenAI in 2022, making it one of the earliest known watermarking proposals for large language models (LLMs). Since then, the field has grown significantly, with many researchers developing their own approaches to the problem.

The author expresses genuine satisfaction at seeing their early work influence how major AI companies are now approaching the challenge of tracking and verifying AI-generated content. LLM watermarking is an increasingly important area of research, as it helps identify text produced by AI systems, which has wide implications for content authenticity, academic integrity, and online trust.

Read original →