Anthropic 在全球范围内为 Claude 生成的文本添加水印
TECH

Anthropic 在全球范围内为 Claude 生成的文本添加水印

53+
Signals

战略概览

  • 01.
    Anthropic 于2026年8月11日宣布,新的 Claude 模型会将一种不可见的统计水印直接嵌入生成的文本中,并在生成的 .svg、.png 和 .jpg 文件上附加签名的 C2PA 来源元数据,主要目的是遵守欧盟《人工智能法案》的透明度要求[1]
  • 02.
    此次部署覆盖了2026年8月2日或之后发布的所有 Claude 模型,适用于 Claude 应用程序、Claude 平台 API、Claude Code、Claude Cowork、Claude Tag 以及在 AWS、Google Cloud 和 Microsoft Foundry 上的云合作伙伴部署——这一措施是全球性的,不仅限于欧盟,且无法选择退出。
  • 03.
    2026年8月12日,即首次公告发布一天后,一名 Anthropic 工程师确认将推出配套的、公众可用的文本检测 API,尽管其定价和访问细节尚未公布。
  • 04.
    Anthropic 自己的文档警告称,该水印可能因大量编辑、改写、翻译、格式转换、重新保存或截图而被削弱或移除,且检测到水印仅意味着 Claude 可能处理过该文本,而非由其创作。

深度分析

一种‘隐形’水印,融入词语选择之中,而非隐藏代码

这种水印并非隐藏的 Unicode 字符或附加在输出末尾的后缀——Anthropic 将其描述为一种在 Claude 生成文本时直接融入其词语选择中的不可察觉信号[1]。这种机制也解释了为何该标记能够随复制粘贴传播,并在轻微编辑后仍能保留——因为底层的词汇和表达方式承载着它——但在大量编辑、改写、翻译或格式转换下会退化,因为原始文本的措辞已所剩无几[7]。该技术基于统计和词元选择,而非一组可见字符,不会改变文本的意义或质量[1][4]。图像和文件则采用另一种更简单的机制:签名的 C2PA 来源元数据,实质上是一张附在 .svg、.png 和 .jpg 输出文件上的数字收据[1]。重新保存文件即可剥离该元数据[7]

一项合规要求,但 Anthropic 选择在全球范围实施

直接动因明确且有时间点:欧盟《人工智能法案》第50条透明度行为准则于2026年8月2日生效,要求生成式 AI 提供商使其合成内容具备机器可检测性,否则将面临高达1500万欧元或全球年营业额3%的罚款[2]。Anthropic 原本可以仅将水印限制在欧盟流量中。但它选择在模型层面于所有运行 Claude 的地区实施——包括 Claude 应用、API、Claude Code、Claude Cowork、Claude Tag 以及云合作伙伴部署——而不是维护区域特定的模型行为[1][2]。与 OpenAI 的对比尤为明显:OpenAI 同样签署了相同的欧盟行为准则,但据报道已搁置一款高精度文本检测器约两年未发布,且尚未为 ChatGPT 推出文本水印,尽管其确实在图像中嵌入了 Google 的 SynthID[3]。Anthropic 的全球、无退出选项的部署,与其说是最低限度的合规,不如说是一种战略定位。

两个争议焦点:代码完整性与披露陷阱

反对声集中在两个不同的风险上。第一个是技术性的:源代码由于严格的语法和标识符限制,不适合承载基于词元的统计水印,而 Prettier 等常见的格式化工具很可能削弱或消除 Claude Code 输出中的水印,从而破坏其作为来源信号的可靠性[4][5]。第二个是法律层面的,据 Stephen Smith 称,后果更为严重——水印本身不包含任何提示词、客户或案件内容,因此并非许多人假设的那种责任来源。他指出的真正风险在于,组织未能跟踪需在申报文件和专业工作成果中披露 AI 使用情况的地域性规则;相比之下,水印只是一个转移注意力的幌子,真正的合规缺口在于此[6]

为何检测到的标记证明力如此之弱——且极易被擦除

Anthropic 自己的文档削弱了对该功能最强烈的主张:检测到的标记仅表明文本可能经过了 Claude 处理,而非由其创作,因为对人类撰写的文本进行校对、翻译或转换仍会产生带标记的输出[1][4]。反之亦然——没有检测到标记并不能证明文本是人类撰写的[1][4],且 Anthropic 自身也警告,在短文本或经过大量编辑的片段中,检测并不可靠[7]。更复杂的是,能够让外部人员独立验证这些声明的检测 API 目前尚未公开,因此目前只有 Anthropic 能确认某段文字是否被标记[4]

历史背景

开源了 SynthID,并将基于词元概率的水印技术集成到 Gemini 模型中,Anthropic 此次部署常被拿来与此比较。
欧盟《人工智能法案》第50条透明度行为准则生效,要求生成式 AI 提供商使其合成的文本、图像、音频和视频输出具备机器可检测性。
通过更新的 Claude 帮助中心文章公开确认,新的 Claude 模型在全球范围内嵌入了隐形文本水印和 C2PA 文件元数据。
一名 Anthropic 工程师在首次公告发布一天后确认,即将推出自助式文本检测 API。

关键关系图

关键玩家
主题

Anthropic 在全球范围内为 Claude 生成的文本添加水印

AN

Anthropic

Built and shipped the watermarking and C2PA metadata system, signed the EU AI Act Article 50(2) Code of Practice on Transparency, and chose to apply the change globally rather than gating it to EU traffic.

EU

European Commission / EU AI Act Article 50

Regulator whose transparency mandate, effective August 2, 2026, is the direct trigger for the rollout; non-compliance risks fines up to EUR 15 million or 3% of global annual turnover.

GO

Google DeepMind (Gemini / SynthID)

Prior mover on AI text and image watermarking since 2024, whose open-sourced SynthID approach is the reference point Anthropic's system is widely compared against.

OP

OpenAI

Also signed the EU Code of Practice but has reportedly held a highly accurate text detector unreleased for roughly two years and has not shipped ChatGPT text watermarking, though it embeds Google's SynthID for images.

PA

Paying Claude users and law firms

Directly affected by the no-opt-out policy - paying users have voiced broad backlash at having no way to turn the feature off, while law firms face pressure to track local rules requiring disclosure of AI use in filings and professional work product.

事实来源

7 条引用
  1. [1] How Claude marks AI-generated content
  2. [2] EU compliance, delivered globally: Anthropic to watermark Claude's output worldwide
  3. [3] Anthropic Claude Text Watermarking and the EU AI Act
  4. [4] Anthropic Claude Invisible Watermarks and C2PA (August 2026)
  5. [5] Anthropic pledges to embed watermarks to help discern AI slop in sop to EU
  6. [6] The Claude watermark is real, the risk is elsewhere
  7. [7] Anthropic to start watermarking Claude-generated text, images

来源文章

Top 5

THE SIGNAL.

Analysts

他认为水印本身并非对企业或专业人士的真实威胁——它不携带任何特权或客户内容,实质性编辑会使信号变得稀薄直至不可靠。在他看来,真正的风险在于组织未能跟踪并遵守要求在申报文件和专业工作成果中通知 AI 使用情况的地方性披露规则。

Stephen Smith
作者,smithstephen.com(法律/科技评论)
The Crowd

🚨HUGE: Claude will now invisibly watermark all AI-generated text, making it detectable even after being copied and pasted. The watermark tagging will apply to all new Claude models, including API outputs, worldwide with no opt-out option.

@@coinbureau385

◻️ AI watermarks are now mandatory in Europe - and almost nobody can verify them yet Article 50 of the EU AI Act kicked in on August 2, 2026. The big AI providers are now required to mark their generated content as machine-readable synthetic output. On August 10 Anthropic...

@@Skybornfx31

Claude will begin digitally watermarking marking AI-generated text and images — Anthropic details how it'll comply with the EU's Artificial Intelligence Act

@@tomshardware18

Claude will watermark generated content, thank you EU

@u/N_P_K3500
Broadcast
Invisible watermarks are coming to Claude's AI-written text

Invisible watermarks are coming to Claude's AI-written text

Claude's Invisible Watermark: What It Can't Prove

Claude's Invisible Watermark: What It Can't Prove

Claude's New AI Models Are Watermarked

Claude's New AI Models Are Watermarked