New Claude Models Will Embed Invisible Watermarks in Generated Text

cryptonews.ruPubblicato 2026-08-11Pubblicato ultima volta 2026-08-11

Introduzione

Anthropic has announced that starting August 2, 2026, new Claude models deployed in the EU will incorporate machine-readable, invisible watermarks into all AI-generated text. This watermark will be woven directly into the model's output, persisting through copy-paste and some editing, without altering readability or meaning. The rule will apply globally across all Claude platforms, including API, apps, and partner services like AWS and Google Cloud. For generated files (e.g., PNG, JPG), a separate, existing system using C2PA metadata signatures will continue to indicate Claude's involvement. Models released before August 2, 2026, including the current Claude Opus 5, are in a transition period and do not yet embed text watermarks, though Anthropic is working to add support. Currently, no public tool exists to detect these text watermarks; Anthropic says it is developing separate detection tools and will release technical details later. The company cautions that even with detection, a watermark does not conclusively prove Claude authored the text, as it is often used to edit or summarize existing content. The timing aligns with the EU AI Act's Article 50, which from August 2, 2026, mandates clear labeling of AI-generated content, with potential fines for non-compliance. A key technical question remains the watermark's resilience to paraphrasing or translation by other AI models, a challenge noted in previous industry experiments.

Anthropic has revealed a previously unpublished detail: starting August 2, 2026, new Claude models launched in the European Union will support machine-readable watermarking of generated text from their release. This refers to future releases—no model has been released after that date yet, the latest currently being Claude Opus 5, released on July 24. This is stated in an article in the Help Center, updated on August 10, 2026. The rule will apply not only in the EU, but wherever Claude is offered—in the API, the Claude app, Claude Code, Claude Cowork, Claude Tag, as well as with partners like AWS, Google Cloud, and Microsoft Foundry.

Text and Files—Two Different Mechanisms

For text, Anthropic will use invisible watermarks that will be woven directly into the model's response itself, without changing its meaning, quality, or readability. Since the mark will become part of the text itself, it will be preserved during copy-paste and, partially, even after editing. This mechanism has no relation to the C2PA standard.

For files—for example, .svg, .png, and .jpg—a different, already operational mechanism applies: Claude attaches signed provenance metadata using the open C2PA (Coalition for Content Provenance and Authenticity) standard. Such a signature indicates that the file was processed by Claude and allows for the detection of tampering with the metadata after the fact. Support for this feature depends on the specific platform.

Old Models Without Text Watermarks

The rule regarding text concerns only new versions of Claude. Models released before August 2, 2026—meaning all currently available models, including Claude Opus 5—are in a transitional period and do not embed text watermarks. Anthropic says it is working on adding support for watermarking for them as well, and the documentation will be updated as readiness is achieved.

Cannot Check Text Yet

Anthropic states that it is working on separate tools that will allow users and third parties to detect text watermarks. The company promises to reveal details of the detection mechanism in technical documentation later—no such tool exists at the moment. Sites like c2pa.org or contentcredentials.org are not suitable for this: they check metadata in files, not text watermarks.

An important nuance: even when detection becomes available, the presence of a mark will not be definitive proof that the author of the text is specifically Claude, since the model is often used for editing, translating, or summarizing others' content. The absence of a mark will also prove nothing.

AI Opinion

From the perspective of macroeconomic regulation, the coincidence of dates does not appear accidental. The European Union, starting August 2, 2026, introduced mandatory requirements of Article 50 of the AI Act regarding the labeling of AI content, and the European Commission had already approved the final clarifications for this norm on July 20, providing for fines of up to 15 million euros or 3% of the company's global turnover for violating transparency. The launch of Claude's text watermark precisely on this day is, apparently, not a technical coincidence but a compliance reaction to a specific regulatory deadline.

The technical aspect that the article does not reveal is the resilience of such watermarks to paraphrasing and translation via another model. Industry experiments with text marks have shown that paraphrasing noticeably reduces detection accuracy, and translating text into another language often removes the mark completely. Whether Claude's invisible mark will remain detectable after processing by a third-party AI is a question whose answer only the company itself currently knows.

Domande pertinenti

QWhat is the new requirement for Claude models launched in the European Union starting August 2, 2026?

AStarting August 2, 2026, new Claude models launched in the European Union will be required to embed machine-readable watermarks into generated text at the time of release. This rule applies wherever Claude is offered, including API, the Claude app, and through partners like AWS and Google Cloud.

QHow does Anthropic's text watermarking for Claude differ from its file watermarking mechanism?

AFor text, Anthropic will use invisible watermarks woven directly into the model's response, which persist through copy-pasting and partially through editing. For files like .png and .jpg, Claude attaches signed metadata of origin using the open C2PA standard, which indicates Claude processed the file and can detect subsequent tampering.

QWill Claude models released before August 2, 2026, have text watermarking?

ANo, Claude models released before August 2, 2026, including the current Claude Opus 5, are in a transitional period and do not embed text watermarks. Anthropic is working to add support for them later.

QCan users currently detect text watermarks in Claude's output?

ANo, users currently cannot detect text watermarks. Anthropic states it is developing separate tools for users and third parties to detect these watermarks, but such tools do not exist yet. Websites like c2pa.org are not suitable as they check file metadata, not text watermarks.

QWhat is the likely reason for Claude's text watermark launch coinciding with August 2, 2026, according to the article's analysis?

AThe launch date likely aligns with compliance to the EU's AI Act. Starting August 2, 2026, the EU enforces mandatory AI content labeling requirements under Article 50, with potential fines for non-compliance. The timing appears to be a regulatory compliance reaction rather than a technical coincidence.

Letture associate

Trading

Spot
活动图片