Anthropic 在官方支持文档中说明 Claude 如何标注 AI 生成内容。该机制旨在提升内容透明度,帮助用户识别由 Claude 生成或修改的文本、图像等输出。具体标注...
行业动态Hacker News 热门(buzzing.cc 中文翻译)
今日 AI 情报摘要
Anthropic 在官方支持文档中说明 Claude 如何标注 AI 生成内容。
该机制旨在提升内容透明度,帮助用户识别由 Claude 生成或修改的文本、图像等输出。 具体标注方式与适用范围以官方文档为准。 🔗 阅读原文 via AIHOT · https://aihot.virxact.com/items/cmsoaprnd0jisrofwjxqn3trv
中文正文 · AI 翻译
Anthropic 已作为生成性 AI 模型和生成性 AI 系统的提供者,签署了欧盟人工智能法案(EU AI Act)第 50(2) 条关于 AI 生成内容透明度的行为准则。本文描述了我们计划如何将这些承诺付诸实践,标记的工作原理,以及其局限性。我们会更新本文,并在更多技术指导可用时发布更详细的说明。
Anthropic 在欧盟人工智能法案关于 AI 生成内容透明度行为准则下的承诺
我们的标记承诺对 Claude 意味着什么:
新模型将从第一天起标记 AI 生成内容。2026 年 8 月 2 日及之后在欧盟推出的 Claude 模型在发布时将支持机器可读标记。生成的文本将带有嵌入式水印,生成的文件在支持的情况下将包含数字签名的来源元数据。
标记在使用 Claude 的任何地方都有效。标记将适用于支持的 Claude 模型在 Claude 平台(API)、Claude、Claude Code、Claude Cowork 及 Claude Tag 上的输出,以及 Claude 在全球提供的任何地方。一些平台或功能可能不支持某些标记类型。
我们会帮助您检测 Claude 的标记。根据准则要求,我们将支持用户和其他第三方检测 Claude 的标记,并将在即将发布的文档中分享详细信息。
Anthropic has signed the EU AI Act's Article 50(2) Code of Practice on Transparency of AI-Generated Content, as a provider of both generative AI models and generative AI systems. This article describes how we’re planning to put those commitments into practice, how marking works, and what its limitations are. We’ll update this article and publish more detailed technical guidance as it becomes available.
Anthropic’s commitments under the EU AI Act’s Code of Practice on Transparency of AI-Generated Content
What our marking commitments mean for Claude:
New models will mark AI-generated content from day one. Claude models launched in the EU on or after August 2, 2026 will support machine-readable marking at launch. Generated text will carry embedded watermarks, and generated files will include digitally signed provenance metadata where supported.
Marking works everywhere you use Claude. Marks will apply to output from supported Claude models across Claude Platform (API), Claude, Claude Code, Claude Cowork, and Claude Tag, and wherever Claude is offered, worldwide. Some platforms or features may not support certain marking types.
We'll help you detect Claude's marks. We'll support users and other third parties to detect Claude’s marks, as the Code requires, and we’ll share details in forthcoming documentation.
Existing models are in progress. The law includes a transition period for Anthropic models launched before August 2, 2026, and we’re working to add marking support for those models as well.
More details about our marking plans are below.
As AI-generated content becomes commonplace, greater transparency and signals about where content comes from can give people useful context about the information they consume. To support transparency and comply with our legal obligations, Anthropic is working to include machine-readable marks in content that Claude generates.
Models. Claude models launched on or after August 2, 2026 support marking at launch. We’re also working to add marking support to Claude models released before that date, and we’ll update this article as that becomes available.
Products. Claude markings cover output from supported models everywhere you use Claude, including Claude Platform (API), Claude, Claude Code, Claude Cowork, and Claude Tag. Embedded watermarks will apply to all generated text. Provenance metadata will apply where Claude supports processing files.
Cloud partners. Embedded watermarks will apply when supported Claude models are accessed through AWS, Google Cloud, or Microsoft Foundry. Signed provenance metadata may not be supported on every platform, depending on the features each platform offers.
Regions. Marking will apply to output from supported models wherever Claude is offered, worldwide.
Claude uses two complementary techniques to mark content generated and processed by Claude: (1) watermarks embedded in text, and (2) signed provenance metadata attached to files.
When a supported Claude model generates text, it weaves an imperceptible watermark directly into the text itself. You won’t see it, and it doesn’t change the meaning, quality, or readability of Claude’s response.
Because the watermark is part of the text, it will travel with the text when it’s copied and pasted elsewhere, and may persist through some editing. Watermarking will be applied at the model level, which means it will be present no matter which Claude product or surface the text comes from.
When Claude generates a supported file type, such as a .svg, .png, or .jpg, it will attach signed provenance metadata. This metadata follows the Coalition for Content Provenance and Authenticity (C2PA) open standard, which is used across the industry to record information about content provenance. If a signed metadata label is present, it signals that a file was processed by Claude and lets you detect whether the file has been tampered with.
We’re also working to enable users and other third parties to detect Claude’s embedded watermarks and provenance metadata. Detection checks whether a piece of text or a file carries a supported Claude mark. If a supported mark is found, it indicates that the content may have been processed by Claude.
We’ll share details on detection mechanisms in forthcoming technical documentation.
Machine-readable marks provide important signals about content, but it’s worth understanding their limitations across all content types.
A detected mark provides a signal that content was processed by Claude, but is not fully conclusive. Detecting a Claude mark tells you that the content may have been processed by Claude. It does not, on its own, confirm the full provenance of the content. For example:
Claude may not be the original author. People often use Claude to proofread, translate, summarize, or convert files. The output can carry a Claude mark even if the underlying ideas, text, or data originated from another source;
The content may have changed after Claude processed it. Marked content may be modified, excerpted, or combined with other material after Claude processed it.
Lack of a detected mark doesn’t mean the content wasn’t AI-generated or processed. Content generated by Claude may not carry a detectable mark if, for example:
It was generated by a model released before marking was supported;
The text has been heavily edited, paraphrased, translated, or mixed into other writing;
The passage is very short, leaving too little text for a reliable signal;
A file’s metadata was stripped through format conversion, re-saving, screenshots, or other means;
It was produced through a platform, feature, or file type where a particular marking type wasn’t supported.
If you deploy Claude in your own product, you should independently assess what Article 50 requires of your products and services. Consistent with our commitments under the EU Code, our goal is to support you in meeting your own transparency obligations, and we'll share technical guidance on our marking and detection approach as it becomes available.
情报判断
Aioga 编辑摘要
Aioga 编辑摘要:Anthropic 在官方支持文档中说明 Claude 如何标注 AI 生成内容。 Aioga 将其归入「行业动态」方向,重点关注它对真实使用和行业竞争的影响。