Claude's new Scarlet Letter watermark is invisible — for now
摘要
Anthropic宣布将为其模型处理(而非仅生成)的内容添加水印,以符合欧盟《人工智能法案》要求。该规定适用于2025年8月2日后发布的所有模型,现有模型需在2026年12月前完成更新。Anthropic确认,未来全球范围内所有新模型都将从上线首日起对AI生成内容嵌入不可见水印,并附带数字签名来源元数据。值得注意的是,其水印覆盖范围超出欧盟最低要求——即使仅
Anthropic has revealed that it will soon watermark content that is processed (not just generated!) by any of its models. In a support article, Anthropic explained that it was rolling out machine-readable watermarks to comply with the European Union’s AI Act, which requires all AI system providers to watermark AI-generated or manipulated audio, image, text, and video outputs. The law applies to any AI model released after August 2 and provides a grace period until December 2026 for providers to update previously released models.
Anthropic confirmed that moving forward, all new models offered globally—not just in the EU—will mark AI-generated content “from day one.” Text outputs will “carry embedded watermarks,” invisible to the user, and other “generated files will include digitally signed provenance metadata where supported,” Anthropic said.
Notably, Anthropic is deploying a "nuke it from orbit" approach, applying the watermarks to all processed content where supported, even though the EU does not require it for cases where an AI system performs "an assistive function for standard editing" (the guidance's own example is grammar correction), or where it doesn't "substantially alter" the user's text or its meaning. A watermark applied at the model level can't tell wholesale generation from a comma fix, so Claude may end up stamping exactly the content the law was written to leave alone. How thoroughly it truly watermarks will not be known until Anthropic releases a detection tool that can be tested. The company said that it plans to eventually share details on how to detect marks in order to offer technical support that the EU’s law requires.
转载信息
评论 (0)
暂无评论,来留下第一条评论吧