Anthropic、Claude の出力にウォーターマーク導入へ批判も
本文の状態
日本語全文を表示中
詳細モードで約5分の本文を読めます。
同じ出来事の情報源
この情報源を基点に整理
The Decoder
Anthropic は Claude の生成テキストに検出可能なウォーターマークを導入したが、John Gruber 氏らは語彙選択の質を損なうと批判し、法務業界では透明性に関する新たな課題が生じている。
AI深層分析を開く2026年8月17日 21:07
AI深層分析
キーポイント
技術的実装と Anthropic の主張
Anthropic は Google の SynthID-Text に基づく手法を採用し、単語選択のランダム性を調整して統計的なパターンを作成するが、可視マークや隠し文字は追加しないと発表している。
John Gruber 氏による品質への批判
Daring Fireball の John Gruber 氏は、ウォーターマーク鍵に基づいて語彙が選択されることで意味の精度が損なわれ、最良の単語ではなく劣る単語の確率が高まると指摘している。
法務業界における透明性の課題
Artificial Lawyer の分析によると、AI 使用を禁止したクライアントや裁判官が skeptical な場合、事実上正確な文書でも AI 由来であることが証明されれば手続きに影響する可能性がある。
Google の研究に対する懐疑論
Gruber 氏は Anthropic が根拠とする Google Deepmind の SynthID 研究における評価指標がテキストの質を正しく測るものではないとし、Gemini の評判低下との関連も示唆している。
ウォーターマークの継承と重複
複雑な契約書ではAI生成部分に付与されたマークが継承され、複数のLLMを使用すると単一の文書内で異なるマークが重なる可能性がある。
重要な引用
When Claude picks between "overcast" and "grey" based on a secret watermark key rather than semantic precision, text quality inevitably suffers.
The markings are harmless, Artificial Lawyer argues, and in most cases neither clients nor courts would object to AI use.
If clients have explicitly banned AI use for their cases, or if a judge is skeptical of it, the AI contribution could be proven even when the submitted document is factually flawless.
Anyone who wants to strip the markings can do so with paraphrasing tools like Declaude.
Source Article
元記事を日本語で読む
本文に関係しない購読案内、埋め込み通知、サイト内プロモーションは除いています。
Anthropic が Claude の出力に導入したテキストウォーターマークは、AI 生成コンテンツの検出を可能にするものとして発表されました。しかし、この手法が語彙選択に影響を与えないかという懸念や、弁護士たちが直面する新たな透明性の課題について、批判の声が上がっています。
Anthropic は同社の Claude のテキストウォーターマーク が、コンテンツの質や創造性、可読性に一切影響を与えないと主張しています。この手法は Google の SynthID-Text に基づくもので、テキスト生成時に単語を選択する際のランダム性を調整し、統計的に検出可能なパターンを埋め込む仕組みです。ただし、目に見えるマークが挿入されるわけではなく、隠れた文字が追加されることもありません。
しかし、著名なブロガーの John Gruber 氏はこれに異議を唱えています。Gruber 氏は Daring Fireball に掲載した長文の投稿で、「同じ意味を持つ同義語であっても、完全に同一の意味を持つものはない」と指摘。Claude が「雲行きが怪しい(overcast)」と「灰色の(grey)」といった単語を選ぶ際、意味の正確さではなく秘密のウォーターマークキーに基づいて判断を下すことで、テキストの質が必然的に低下すると述べています。Gruber 氏によれば、このシステムは時には不適切な単語の確率を上げ、最適な単語の確率を下げることもあるといいます。Anthropic が「その違いは知覚できない」と主張している点については、Gruber 氏は単に誤りだと断じています。
Gruber 氏は 2002 年から影響力のあるテックブログ『Daring Fireball』を運営しており、また現在チャットボットなどで広く利用されているマークダウン(Markdown)の共同開発者でもあります。
Google DeepMind が Nature に発表した SynthID に関する研究は、Anthropic も引用しているが、Gruber 氏を納得させるには至らなかった。同氏は、そこで測定された「高評価・低評価」の比率はテキスト品質を測る有効な指標ではないと主張する。「バナナ」という単語を「パイナップル」と間違えたからといって、誰もチャットボットに低評価をつけるわけではないからだ。Gruber 氏によれば、Gemini が Claude や ChatGPT よりも劣っていると見なされている評判は、SynthID の機能が既に稼働していることが一因となっている可能性もある。
法律事務所は透明性に関する質問への備えが必要だ
法律専門誌『Artificial Lawyer』は、Anthropic のウォーターマークが法律事務所にどのような意味を持つのかを検証した。同誌の結論は概ね楽観的だ。同誌は、これらのマーキングは無害であり、ほとんどの場合、クライアントも裁判所も AI の利用に異議を唱えないと指摘している。むしろ分析によると、一部のクライアントからはすでに積極的にその利用を求めているという。
ただし、『Artificial Lawyer』の評価によれば、特定の状況では話が変わってくる。クライアントが自らの事件における AI の利用を明確に禁止している場合や、裁判官が AI に懐疑的な態度を示す場合は、提出された文書の内容が事実上完璧であっても、AI の関与が証明されてしまう可能性がある。それが訴訟手続きや裁判の行方にどのような影響を与えるかは注目される点だ。
また、この透かしはテキストとともに移動する点も『Artificial Lawyer』が指摘しています。複雑な契約書には、AI が生成した部分にのみ透かしが含まれ、残りは人間が作成しているというケースもあり得ます。過去のテンプレートに基づいて構築される将来の契約では、これらのマークが継承されます。複数の大規模言語モデル(LLM)を使用する場合、1 つの文書内で異なる透かしが重なり合う可能性さえあります。
料金交渉の行方も変化する可能性があります。クライアントが AI によって作業が容易になったことを理由に値引きを求めた場合、AI の貢献度が原則として検証可能になると、同誌は指摘しています。
Anthropic 自身は、事実記述が多い箇所では透かしが疎になる傾向があると説明しています。これは語彙の選択肢が少ないためです。正確性が求められる法律文書においては重要な注意点ですが、現時点でこの分野に関する実証研究はまだ行われていません。
これらのマークを除去したい場合は、Declaude といった言い換えツールを使用できます。開発者のジェームズ・パドルシー氏は、このシステムの基盤となっている EU の規制が恣意的だと批判しています。彼は、このシステムは意図的な回避対策としてはほとんど機能せず、主に一般ユーザーにのみ影響を与えていると述べています。
Anthropic は EU AI 法への対応として透かし機能を導入していますが、地域ごとの制限ができないためグローバルに展開しています。8 月以降にリリースされた Claude モデルはすべてこのマークに対応しており、旧モデルについても今後数ヶ月のうちに順次対応が追加される予定です。
AI ニュースを過剰な hype(過熱)なしに – 人間が厳選したニュース
広告なしで記事を読みたい方、毎週届く AI ニュースレター、年に6回発行する独占の「AI Radar」フロンティアレポート、アーカイブへのフルアクセス、そしてコメント欄の利用権をご希望の方は、THE DECODER にぜひご登録ください。
原文を表示
Anthropic's text watermarking for Claude is supposed to make AI-generated content detectable. But critics doubt that word choice stays unaffected, and lawyers are facing new transparency headaches.
Anthropic insists that text watermarking for Claude has no effect on content, creativity, or readability. The method is based on Google's SynthID-Text approach and tweaks the randomness source for word selection during text generation, creating a statistically detectable pattern. No visible marks are inserted. No hidden characters are added.
Blogger John Gruber disagrees. In a lengthy post on Daring Fireball, he argues that no two synonyms carry the exact same meaning. When Claude picks between "overcast" and "grey" based on a secret watermark key rather than semantic precision, text quality inevitably suffers. The system sometimes boosts the probability of a worse word and lowers it for the best one, Gruber writes. Anthropic's claim that the difference is "imperceptible" is simply wrong, he says.
Gruber has run the influential tech blog Daring Fireball since 2002. He also co-created Markdown, the formatting language now widely used by chatbots.
Even the SynthID study that Google Deepmind published in Nature, which Anthropic cites, doesn't convince Gruber. The thumbs-up/down rates measured there aren't a valid gauge of text quality, he argues. Nobody gives a chatbot a thumbs-down because it wrote "bananas" instead of "pineapples." Gruber suggests that Gemini's reputation as weaker than Claude and ChatGPT could partly stem from SynthID already being active.
Law firms need to prepare for transparency questions
Legal trade publication Artificial Lawyer examined what Anthropic's watermarks mean for law firms. The publication's verdict is mostly relaxed. The markings are harmless, Artificial Lawyer argues, and in most cases neither clients nor courts would object to AI use. Some clients are even actively requesting it now, the analysis notes.
Things get tricky in specific situations, though, according to Artificial Lawyer's assessment. If clients have explicitly banned AI use for their cases, or if a judge is skeptical of it, the AI contribution could be proven even when the submitted document is factually flawless. That could have consequences for how proceedings or a trial play out.
The watermarks also travel with the text, Artificial Lawyer points out. A complex contract might contain clauses that are AI-marked while a human wrote the rest. Future contracts built on older templates with AI-generated sections carry those markings forward. When multiple LLMs are used, different watermarks could even overlap in a single document.
Fee negotiations could shift, too. If clients demand discounts because AI made the work easier, the AI share becomes verifiable in principle, the publication notes.
Anthropic itself says that watermarking is sparser in fact-heavy passages because fewer word alternatives exist. For legal texts that depend on precision, that's a relevant caveat, even though there aren't any empirical studies on the topic yet.
Anyone who wants to strip the markings can do so with paraphrasing tools like Declaude. Its developer James Padolsey criticizes the underlying EU regulation as arbitrary. The system mostly hits ordinary users while doing little against deliberate circumvention, he says.
Anthropic is rolling out watermarking to comply with the EU AI Act but applies it globally because it can't limit the feature by region. All Claude models released after August 2 support the marking. Older models will be retrofitted in the coming months.
AI News Without the Hype – Curated by Humans
Subscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive "AI Radar" frontier report six times a year, full archive access, and access to our comment section.
関連記事
今日のまとめ
AIデイリーブリーフで今日の重要ニュースをまとめ読み