Skip to main content
Newsletter IconGoogle IconFacebook IconX IconInstagram IconYouTube Icon

Anthropic is adding imperceptible, model-level watermarks to Claude's AI-generated text

Thanks to the European Union's Artificial Intelligence Act (AIA), Anthropic is embedding 'imperceptible' watermarks in Claude's AI-generated text.

Anthropic is adding imperceptible, model-level watermarks to Claude's AI-generated text
Facebook IconX IconReddit Icon
Comments
Senior Editor
Published
1 minute & 15 seconds read time
TL;DR: Anthropic will embed imperceptible, model-level watermarks into Claude's AI-generated text that persist through copying and some editing, allowing detection that text was processed by Claude though not fully conclusive; image identification instead uses signed provenance metadata, which can be lost if images are screenshot.
Voice: Kosta Andreadis
0:00 / 2:44
Use left and right arrow keys to seek audio.

AI-generated text, images, and video have all reached that level where it's becoming increasingly difficult to tell the difference between what's human-generated versus stuff spat out by hardware sitting in a data center somewhere. This is why the European Union's Artificial Intelligence Act (AIA) requires AI companies to implement identification measures to ensure there's a simple check, while offering a Code of Practice.

Anthropic is adding imperceptible, model-level watermarks to Claude's AI-generated text 1

And on that note, Anthropic has posted a new article on its Claude support portal titled: How Claude marks AI-generated content. Here it details how Claude will begin to embed watermarks in AI-generated text, which the company describes as imperceptible and injected directly into the text itself. It will reportedly persist even with some editing, meaning that for those who do some light editing and reword Claude-generated text, the result will still include the watermark identifying it as AI-generated content.

Popular Now: Redditor buys an RTX 5070 Ti for $1,099 from Walmart, receives four graphics cards instead

"Because the watermark is part of the text, it will travel with the text when it's copied and pasted elsewhere, and may persist through some editing," Anthropic confirms. "Watermarking will be applied at the model level, which means it will be present no matter which Claude product or surface the text comes from."

Anthropic notes that it's also working to enable Claude users and third parties to detect the embedded watermarks, with details on how this will work still to come. Anthropic notes that a positive detection means that it will mark text as being "processed by Claude," but it won't be "fully conclusive." One reason is that someone using Claude to proofread or make changes to a document doesn't necessarily mean Claude is the original author.

Frequently Asked Questions

Open a question for an answer from TweakTown's coverage of this news, or ask your own below.

Question #1

How does Anthropic embed the imperceptible watermark into Claude's AI-generated text?

Anthropic injects the imperceptible watermark directly into Claude's generated text at the model level, so the watermark is part of the text itself and will be present across any Claude product or surface. Because it is embedded in the text, the watermark can travel when the text is copied and pasted and may persist through some light editing. Anthropic is also working to enable detection tools that will mark positively detected text as having been processed by Claude, though such detections will not be fully conclusive.
Question #2

Which Claude products or interfaces will include the watermark in generated text?

Watermarking will be applied at the model level, so it will be present no matter which Claude product or surface the text comes from.
Question #3

How will third parties or users be able to detect whether text was processed by Claude?

Anthropic is embedding imperceptible, model-level watermarks directly into Claude-generated text and says it is working to enable Claude users and third parties to detect those embedded watermarks, with details on detection to come. A positive detection will mark text as being "processed by Claude," but Anthropic cautions that such a result will not be fully conclusive because use of Claude for proofreading or edits does not prove original authorship.
Question #4

What does a positive watermark detection indicate about authorship or provenance?

A positive watermark detection indicates the text was processed by Claude and is marked as having been handled by that model. Anthropic cautions that this is not fully conclusive evidence of original authorship or provenance, since someone could have used Claude only to proofread or edit an existing document.

Have a question that isn't listed here? Ask below, and TweakBot will answer it.

Interestingly, when it comes to AI-generated images, Anthropic's Artificial Intelligence Act (AIA) compliance doesn't actually embed a watermark in the .svg, .png, or .jpg images you see, but rather via attached signed provenance metadata. Basically, the metadata in the file will include the confirmation that it was processed by Claude. When it comes to copying an AI-generated image via a screenshot and then saving it manually, this digital certificate of sorts is lost, so it's not as foolproof as embedding the watermark in generated text at the model level.

Photo of the codeMeter Claude Code Usage Monitor

Best Deals: codeMeter Claude Code Usage Monitor

Prices last scanned 17 hours and 5 minutes ago

* Prices may be inaccurate. As an Amazon Associate, we earn from qualifying purchases. We earn affiliate commissions from Newegg and PCCG sales.

Join Our Newsletter

Join the TweakTown Newsletter for daily tech updates delivered to your inbox.

See previous giveaways.

News Source:support.claude.com

Comments

About the author

Senior Editor

Kosta is a veteran gaming journalist that cut his teeth on well-respected Aussie publications like PC PowerPlay and HYPER back when articles were printed on paper. A lifelong gamer since the 8-bit Nintendo era, it was the CD-ROM-powered 90s that cemented his love for all things games and technology. From point-and-click adventure games to RTS games with full-motion video cut-scenes and FPS titles referred to as Doom clones. Genres he still loves to this day. Kosta is also a musician, releasing dreamy electronic jams under the name Kbit.

Stay Updated

Follow TweakTown for breaking tech news, reviews, and daily updates.

Follow TweakTown on GoogleAdd TweakTown as a preferred source on Google
Newsletter Subscription