Anthropic、EU規制対応のためClaude生成文に世界規模で透かし埋め込み
本文の状態
日本語全文を表示中
詳細モードで約5分の本文を読めます。
同じ出来事の情報源
この情報源を基点に整理
Euronews Next AI
Anthropic は EU の AI 法対応として、Claude が生成するテキストやファイルに不可視透かしとメタデータを付与する機能を全世界で導入すると発表した。
Continue in AI NEW LAB
このニュースを、実務の判断につなげる
AI NEW LABで、試したことや先に確認したい条件を共有できます。まずはログインなしで読めます。
AI NEW LABで論点を見るAI深層分析を開く2026年8月11日 22:01
AI深層分析
キーポイント
EU 規制への対応とグローバル展開
Anthropic は EU の AI 法(Article 50)の透明性要件を満たすため、Claude の生成物に透かしを埋め込む方針を示し、この措置は EU 域外を含む全世界で適用される。
2 つの技術的アプローチ
同社はテキストへの不可視パターン埋め込みと、画像・ドキュメントへのデジタル署名付きメタデータ付与という 2 つの手法を併用し、改ざん検知や生成元識別を可能にする。
適用範囲と対象サービス
この透かし機能は、消費者向けアプリ、API(Claude Platform)、開発ツール(Claude Code, Cowork, Tag)および AWS、Google Cloud、Microsoft Foundry を介したアクセスを含む全サービスに及ぶ。
メタデータの限界
同社はフォーマット変換やスクリーンショットなどの処理によってメタデータが失われる可能性を明示しており、完全な追跡保証は限定的であることを認めている。
EU AI Act に基づく義務と罰則
EU AI Act の第50条により、2026年8月2日以降、生成AIは合成コンテンツにマーキングを行う義務が生じる。違反した場合、最大1,500万ユーロまたは全世界年間売上の3%のいずれか高い方の罰金が科される。
重要な引用
"Claude models launched on or after August 2, 2026 support marking at launch. We're also working to add marking support to Claude models released before that date"
"The watermark will either be weaved in as 'an imperceptible watermark directly into the text itself' or as metadata"
"Anthropic said a file's marking can be stripped through format conversion, re-saving, screenshots, or other similar processes"
Regulators expect providers to rely on a multi-layered marking strategy, combining digitally signed metadata, imperceptible watermarking and, in some cases, fingerprinting or logging as a fallback, since the Code of Practice makes clear that no single marking technique is sufficient on its own.
編集コメントを表示
編集コメント
EU の AI 法が単なる地域規制にとどまらず、グローバルな業界標準を形成する転換点となった事例である。同社はメタデータの脆弱性も正直に開示しており、技術的な限界と法的要件のバランスを取る姿勢が見て取れる。
Source Article
元記事を日本語で読む
本文に関係しない購読案内、埋め込み通知、サイト内プロモーションは除いています。
Anthropic は、生成したテキストに目に見えない透かしを埋め込む措置を開始します。これは、人工知能(AI)の透明性に関する欧州連合(EU)の新規規則への対応を目的としたものです。
「2026 年 8 月 2 日以降にリリースされる Claude モデルは、起動時に透かし付けをサポートしています。また、それ以前にリリースされたモデルについても同様のサポートを追加する作業を進めています」と、サンフランシスコに拠点を置く AI 大手は声明で述べています。
この新しい透かし付けポリシーには、Claude が提供するすべてのサービスが含まれます。一般消費者向けアプリや開発者向けの Claude Platform(API)、Claude Code、Claude Cowork、Claude Tag、さらに AWS、Google Cloud、Microsoft Foundry を通じてアクセスできる Claude のバージョンも対象です。
透かしは、「テキストそのものに直接織り込まれる不可視の透かし」として埋め込むか、メタデータとして付与されます。これにより、「Claude が .svg、.png、.jpg などの対応ファイルタイプを生成する際、署名付きのプロベナンス(出所)メタデータを添付し、そのファイルが Claude によって処理されたことを示すと同時に、改ざんの有無を検知できるようにします」。
Anthropic は、この透かしは EU のみならず、Claude が提供されるすべての地域に適用されると説明しています。
透かし付けの適用範囲は欧州に限られません。Anthropic は、Claude が世界中で提供されている場所すべてにこの透かしを適用すると明言しました。これは、EU のコンプライアンス基準が AI に関するグローバルな標準設定にも波及効果をもたらす好例と言えます。
How the watermark works
Anthropic は2 つの異なる手法を採用しています。1 つ目は、生成されたテキストに直接、人間には認識できないパターンを埋め込む技術です。このパターンは読者には見えないものの機械には検出可能で、テキストがコピー&ペーストされて他の場所へ移動してもその痕跡と共に存在し続けます。
同社によると、この手法によって Claude の回答の意味や品質、可読性が損なわれることはありません。
2 つ目の手法はファイルを対象としています。これは写真の EXIF データに似たデジタルな足跡のようなもので、隠すためではなく、検証可能であることを目的に設計されています。
実際にこれを確認するには、標準的な「プロパティ」メニューではなく、C2PA に対応したツールか Anthropic が計画している検出システムが必要です。また、この手法は Claude という名前を特定するものというよりは、「AI システムを経由して生成されたファイルである」というフラグを立てるものです。
同社はまた、メタデータが失われる可能性があることを明確にしています。
Anthropic によれば、ファイルのマークはフォーマット変換や再保存、スクリーンショット撮影など、類似のプロセスを通じて剥奪される可能性があります。つまり、Claude のマークを付与していた文書や画像も、編集や変換が行われた後には痕跡が完全に消えている可能性があるということです。
なぜ今なのか
このタイミングは、EU の AI 法に直接起因しています。
同法では第50条において、ディープフェイクなどの合成コンテンツを生成する AI システムに対し、その出力を人工的に生成されたものとして表示することを義務付けています。
規制当局は、単一の標識技術だけでは不十分であることが運用規範で明確に示されているため、デジタル署名付きメタデータ、不可視の透かし、そして場合によっては指紋認証やログ記録をフォールバックとして組み合わせた多層型の標識戦略に依存することを提供者に求めています。
第50条に基づく透明性義務は2026年8月2日に発効しました。これに違反した場合、罰則として1,500万ユーロまたは全世界の年間総売上の3%(いずれか高い方)が科される可能性があります。
個人は依然として痕跡を消せるのか?
Anthropic は明確に述べています。検出された透かしは「証拠」ではなく「シグナル」に過ぎない、と。
テキストにマークが表示されたとしても、それは単にそのテキストが過去に Claude を経由した可能性を示すだけであり、そのテキストがどのように作成されたかという完全な履歴を証明するものではありません。
利用者は、他で生成された資料の校閲、翻訳、要約のためにアシスタントをよく使用します。そのため、Claude が元から執筆したわけではない文章にも透かしが表示されることがあります。
逆もまた然りです。大幅な編集や翻訳、非常に短い文章、あるいはマーク機能に対応していない旧モデルの使用などは、Claude による正当な出力が検出されない原因となり得ます。
これは Anthropic だけの問題ではありません。Google DeepMind はすでに AI が生成したテキスト、画像、音声に标记するための「SynthID」を開発済みです。また OpenAI も透かし手法について議論を行っていますが、広範な導入についてはより慎重な姿勢を示しています。
原文を表示
Anthropic will begin embedding invisible watermarks into text generated by its Claude large language model and chatbot in a move designed to satisfy new European Union rules on AI transparency.
"Claude models launched on or after August 2, 2026 support marking at launch. We’re also working to add marking support to Claude models released before that date," the San Francisco based AI giant said in a statement.
The new marking policy includes all of the services provided by Claude, including its consumer app, the developer-facing Claude Platform (API), Claude Code, Claude Cowork and Claude Tag, as well as versions of Claude accessed through AWS, Google Cloud and Microsoft Foundry.
The watermark will either be weaved in as "an imperceptible watermark directly into the text itself" or as metadata, so that when "Claude generates a supported file type, such as a .svg, .png, or .jpg, it will attach signed provenance metadata... [that] signals that a file was processed by Claude and lets you detect whether the file has been tampered with."
Anthropic said the watermark would apply to every region where Claude is offered, not just the EU.
The marking will not be limited to Europe. Anthropic said the watermark would apply everywhere Claude is offered worldwide, another example of how the EU's standard of compliance has ripple effects on the global standards being set for AI.
How the watermark works
Anthropic uses two separate techniques. The first includes inserting an imperceptible pattern directly into generated text, invisible to readers but detectable by machines and able to travel with the text even after it has been copied and pasted elsewhere.
According to the company, it doesn't change the meaning, quality or readability of Claude's response.
The second technique applies to files. It functions more like a digital paper trail, similar to EXIF data on a photograph, and is designed to be checkable rather than concealed.
Seeing it in practice would likely require a C2PA-aware tool or Anthropic's planned detection system, rather than a standard "properties" menu, and it would not necessarily identify Claude by name so much as flag that the file passed through an AI system.
The company has also been explicit that the metadata can be lost.
Anthropic said a file's marking can be stripped through format conversion, re-saving, screenshots, or other similar processes, meaning a document or image that once carried a Claude mark may no longer show any trace of it after being edited or converted.
Why now
The timing traces directly to the EU's AI Act.
The act requires that AI systems creating synthetic content, such as deepfakes, mark their outputs as artificially generated under Article 50.
Regulators expect providers to rely on a multi-layered marking strategy, combining digitally signed metadata, imperceptible watermarking and, in some cases, fingerprinting or logging as a fallback, since the Code of Practice makes clear that no single marking technique is sufficient on its own.
The transparency obligations under Article 50 took effect on 2 August 2026. Non-compliance can trigger fines of up to €15 million or 3% of total global annual turnover, whichever is higher.
Can people still cover their tracks?
Anthropic has been explicit that a detected watermark is a signal, not proof.
A mark showing up in a piece of text confirms only that it may have passed through Claude at some point, not the full history of how that text came to be.
People often use the assistant to proofread, translate or summarise material that originated elsewhere, so a watermark can appear on text Claude did not originally write.
The reverse is also true: heavy editing, translation, very short passages, or use of an older model without marking support can all mean genuine Claude output goes undetected.
Anthropic is not alone in this. Google DeepMind has already developed SynthID for marking AI-generated text, images and audio, while OpenAI has discussed watermarking approaches but has moved more slowly to deploy them broadly.
関連記事
今日のまとめ
AIデイリーブリーフで今日の重要ニュースをまとめ読み