你看不到它。但它在那里。
8月17日,Anthropic宣布:未来版本的Claude模型将引入隐形文本水印技术,以满足欧盟AI法案(EU AI Act)的合规要求。
关键是这个水印的实现方式——不是很多人想象的"在文本里插入看不见的隐藏字符",也不是额外的标签或标记,更不包含用户身份信息。它是基于模型词元选择中的统计模式——模型在选择下一个词的时候,会按照一种只有Anthropic知道的统计规律来选择,这样生成的文本看起来完全正常,但Anthropic可以通过统计分析检测出这是不是Claude生成的。
为什么这很重要?
首先,这标志着欧盟AI法案的执行正在进入实质阶段。
欧盟AI法案是全球第一部综合性的AI监管法律,高风险AI系统条款将在2026年8月全面适用。其中一个核心要求就是:AI生成的内容必须能够被识别——不管是文本、图片、视频还是音频,用户有权知道自己看到的内容是不是AI生成的。
Anthropic不是第一个做水印的公司,但它是第一个明确说明水印技术细节、而且承诺不包含用户身份信息的头部大模型厂商。更重要的是,Anthropic在公告里明确说:其他主要模型开发商也已签署同一《行为准则》,将同步实施水印方案。
这意味着什么?意味着以后GPT生成的内容、Claude生成的内容、Gemini生成的内容——都会带上各自的"统计指纹"。你不需要在文本里加什么"AI生成"的标签,不需要靠检测工具瞎猜,通过统计分析就能准确溯源。
这对于打击AI深度伪造、AI造谣、AI诈骗当然是好事——以后假新闻、假视频、假声音,一检测就知道是谁家模型生成的,责任主体清晰了很多。但另一方面,这也引发了一些关于隐私和言论自由的担忧:如果所有AI生成的内容都能被溯源,那人们用AI辅助写作、创作、表达的时候,会不会有心理负担?会不会导致自我审查?
AI内容的"出生证明"时代
隐形水印技术的普及,意味着AI内容正式进入了"持证上岗"的时代。
以前的互联网是"默认匿名"的——你在网上发的内容,除非你主动署名,否则很难溯源到具体的人或工具。以后的AI内容是"默认可溯源"的——每一段AI生成的文本、每一张AI生成的图片、每一段AI生成的视频,都会带上模型厂商的水印,就像产品的出厂铭牌一样。
这个变化是好是坏?可能没有简单的答案。
从积极的方面看,这能有效遏制AI滥用——深度伪造、AI造谣、AI写的垃圾邮件、AI生成的诈骗内容,这些问题会得到一定程度的缓解。普通用户识别AI内容的门槛会大大降低,不会再被AI生成的假信息骗得团团转。
从担忧的方面看,这可能会限制AI工具的合理使用——比如一个作者用AI辅助构思、打草稿、润色文字,但最终作品是自己完成的,这算不算AI生成?要不要带水印?检测工具会不会"误伤"?这些问题还需要更细致的规则来界定。
但不管你喜不喜欢,AI内容溯源的时代已经来了。就像互联网从"匿名"走向"实名"是大势所趋一样,AI内容从"无法分辨"走向"可以溯源",也是不可逆转的趋势。Anthropic这次的公告,只是这个趋势里的第一个正式脚印。
明天见。
You can't see it. But it's there.
On August 17, Anthropic announced: future versions of Claude will include invisible text watermarking technology to meet EU AI Act compliance requirements.
The key detail is how this watermark works—it's not "inserting invisible hidden characters into text" as many imagine, nor extra tags or markings, and it does not contain user identity information. It's based on statistical patterns in the model's token selection—when the model picks the next token, it follows a statistical pattern known only to Anthropic. The resulting text looks completely normal, but Anthropic can detect via statistical analysis whether it was generated by Claude.
Why This Matters
First, it signals that EU AI Act enforcement is entering a substantive phase.
The EU AI Act is the world's first comprehensive AI regulation, with high-risk AI system provisions taking full effect in August 2026. One core requirement: AI-generated content must be identifiable—whether text, image, video, or audio, users have the right to know if what they're seeing is AI-generated.
Anthropic isn't the first company to do watermarking, but it's the first major LLM vendor to clearly explain the technical details and commit to not including user identity information. More importantly, Anthropic explicitly stated: other major model developers have signed the same Code of Practice and will implement watermarking in lockstep.
What does that mean? It means content generated by GPT, Claude, Gemini—all will carry their own "statistical fingerprints." You won't need "AI-generated" labels or guesswork from detection tools; statistical analysis can accurately trace provenance.
This is certainly positive for combating deepfakes, AI disinformation, and AI scams—fake news, fake videos, and fake voices can be traced back to the model vendor, clarifying liability. But it also raises privacy and free speech concerns: if all AI-generated content is traceable, will people feel inhibited when using AI to assist with writing, creativity, or expression? Will it lead to self-censorship?
The Era of AI Content "Birth Certificates"
The spread of invisible watermarking means AI content officially enters an era of "licensed to operate."
The old internet was "anonymous by default"—content you posted online was hard to trace back to a specific person or tool unless you actively signed it. Future AI content will be "traceable by default"—every AI-generated text, image, and video will carry the model vendor's watermark, like a factory nameplate on a product.
Is this change good or bad? There's probably no simple answer.
On the positive side, it can effectively curb AI misuse—deepfakes, AI disinformation, AI-written spam, AI-generated scam content will see some relief. Regular users will have a much lower barrier to identifying AI content and won't be fooled by AI-generated misinformation.
On the concern side, it may limit legitimate uses of AI tools—for example, an author who uses AI to brainstorm, draft, and polish but whose final work is original—is that considered AI-generated? Does it need a watermark? Will detection tools produce false positives? These questions will need more nuanced这些问题将需要更细致的规则来界定。
But whether you like it or not, the era of AI content provenance is here. Just as the internet moved from "anonymous" to "identified" as an inevitable trend, AI content moving from "indistinguishable" to "traceable" is also irreversible. Anthropic's announcement is just the first formal footprint in this trend.
See you tomorrow.
Anthropic, invisible watermark, EU AI Act, AI regulation, content provenance, token statistics
Sources · 信源 Sources
本文基于 Dawn Vision 认知引擎处理的 6 个源信号生成,经编辑部人工审核。素材来源:InfoQ。
Generated by the Dawn Vision cognitive engine processing 6 source signals, with human editorial review. Source: InfoQ.