Earlier this month, Anthropic announced that it was adding invisible text watermarking to Claude outputs. This announcement got a lot of attention.
At the same time the European Commission announced that other firms, including Black Forest Labs and Open AI have also committed to taking steps to mark AI-generated outputs.
Because of this, there's been a lot of interest in understanding:
- How AI text watermarking works
- Whether AI text watermarking can be evaded or erased
Here's an in-depth educational resource I developed that answers both questions.
The resource also highlights one potential unexpected benefit of AI text watermarking. We might be able to better answer the question: 'How much human input went into this content?"
[link] [comments]