OpenAI、顧客プライバシー保護を強化し Anthropic を上回る狙い
本文の状態
日本語全文を表示中
詳細モードで約5分の本文を読めます。
同じ出来事の情報源
この情報源を基点に整理
TechCrunch AI
OpenAI は競合の Anthropic を上回るため、顧客データを一切保持しない「Private Safety Processing」という新サービスを発表し、複数セッションにわたる悪用検知を実現する。
AI深層分析を開く2026年8月20日 07:27
AI深層分析
キーポイント
新機能「Private Safety Processing」の発表
OpenAI は競合他社を凌駕するため、顧客データを一切保持せずに潜在的な悪用を検知する自動化システム「Private Safety Processing」を選定顧客向けにプレビューした。
Anthropic のデータ保持方針との対比
この発表は Anthropic が「対象モデル」において 30 日間のデータ保持を認める方針と明確に対立しており、機密データを扱う企業からの懸念に応える形となっている。
ゼロデータ保持(ZDR)の範囲拡大
OpenAI は既存の「ゼロデータ保持」ポリシーを拡張し、単一セッションだけでなく複数会話にわたる入出力を評価する長期的な安全監視機能としてこの技術を位置づけている。
エージェントによる自動検知の実装
監視は人間を介さずエージェントが行い、悪意ある行為者がサイバー攻撃用のマルウェア作成など複数のセッションにわたってリクエストを分散させるケースも検出可能とする。
複数セッションにわたる悪意ある利用の検出機能
新技術は複数のセッションにまたがる攻撃的な使用を検出し、人間のレビューを介さずにユーザーの会話内容を分析する。
重要な引用
OpenAI just announced a privacy-centric safety approach to monitoring for misuse.
This is an automated system that watches for potential abuse while simultaneously retaining none of the customer's data.
The policy, which has aggravated some customers, enables the AI lab to keep user data (all of their sessions — and the conversations therein) for a period of 30 days.
"Private Safety Processing can analyze those multiple conversations for signs of abuse without human review of a user's conversations."
編集コメントを表示
編集コメント
OpenAI が競合のデータ保持方針を逆手に取り、プライバシー重視の安全監視機能を強化した点は業界標準の再定義につながる可能性がある。企業側は自社のコンプライアンス要件と照らし合わせ、プロバイダー選定における新たな基準としてこの動向を注視する必要がある。
Source Article
元記事を日本語で読む
本文に関係しない購読案内、埋め込み通知、サイト内プロモーションは除いています。
AI モデルの能力が向上するにつれ、その悪用されるリスクも高まっています。これに伴い、こうした悪用を防ぐための安全対策を求める声も大きくなっています。現在、AI 企業は、企業の顧客プライバシーを尊重しつつ、同時に利用状況に問題がないか監視するという、微妙なバランスの上に立っています。
競合他社である Anthropic を上回る機会と捉えた OpenAI は、悪用を検知するためのプライバシー重視の安全アプローチを発表しました。同社は現在、「Private Safety Processing(プライベート・セーフティ・プロセッシング)」と呼ばれる新サービスを特定の顧客向けにプレビュー中です。これは、顧客データを一切保持することなく、潜在的な悪用を自動的に監視するシステムです。
この仕組みは、Anthropic が最近発表したデータ保持ポリシーとは明らかに相反するものです。同社が「対象モデル」と呼ぶものについては、ユーザーの全セッションとその会話内容を含むデータを 30 日間保持することを可能にするこのポリシーは、一部の顧客から批判を招いています。OpenAI によると、「対象モデル」には、Mythos クラスのすべてのモデルと、同様の能力を持つ将来のモデルが含まれます。
この方針は7月に発表されたもので、安全性を確保し、不適切な行為を特定・分析するために設計されました。しかし、大量の機密データを扱い、AI ラボがそのデータを保管したり検査したりすることを望まない一部の企業からは、懸念の声が上がっています。
OpenAI は他の多くの AI 企業と同様に、「ゼロデータ保持(Zero Data Retention: ZDR)」という方針を遵守することで、顧客に一定のプライバシー保護を提供しています。ZDR では、OpenAI API 内のエージェントがセッションごとに不正利用を検知します。これにより、企業はデータを保有することなく、人間の介入なしで悪意のある活動をスキャンすることが可能になります。なお、Anthropic も「カバー対象モデル(covered models)」である Fable のような例外を除き、基本的に ZDR を遵守しています。
OpenAI によると、「プライベートセーフティプロセッシング(Private Safety Processing)」は、ZDR の範囲を拡大する新技術です。これは単一の会話だけでなく、複数の会話の入力と出力を評価する「長期にわたる安全性モニタリング」の一種として説明されています。監視自体はエージェントによって行われ、トリガーが作動した場合はセッション間を跨いでやり取りを捕捉し、潜在的な悪用の兆候を検出します。
新しい技術により、OpenAI は複数のセッションにわたって行われる AI の悪用を検出できるようになったと、同社のスポークスマンは TechCrunch に語った。例えばサイバー攻撃用のマルウェアを仕掛けることを狙う悪意のあるアクターは、検知を回避するためにリクエストを分散させる可能性がある。OpenAI によると、「プライベート・セーフティ・プロセッシング(Private Safety Processing)」を使えば、ユーザーの会話内容を人間がレビューすることなく、複数の対話から悪用の兆候を検出できるという。
システムがトリガーされた場合、同社は「特定の活動タイプを警告する、限定的に定義されたシグナル」を OpenAI に送信すると説明している。このシグナルに基づき、OpenAI は「是正措置が必要かどうか」を判断する。必要と判断されれば、OpenAI は顧客に連絡して状況の追加情報を求めたり、問題解決のために協力したりする。その際、顧客は自らの判断でデータを OpenAI と共有することも可能だとスポークスマンは述べた。
一方、Anthropic は顧客データの人間によるレビューが行われる可能性はあるが、それは「承認された少数のレビュアーのみがアクセスできる管理された経路」を通じてのみ行われると指摘している。同社によると、すべてのレビューセッションは「レビュアーが抑制したり改ざんしたりできない不変ログに記録される」という。
OpenAI と Anthropic の企業間競争は現在、緊張感に満ちています。両社は互いに優位に立つためのあらゆる機会を模索しており、その状況は熾烈です。
直近の報道では、OpenAI の第 2 四半期(Q2)の成長率が Anthropic を下回ったことが示されました。一方、Anthropic の年間収益化率(annualized revenue run rate)は現在、約 650 億ドルに達していると報じられています。
Anthropic の投資家からは、同社が将来 2 兆ドル規模で株式公開(IPO)を果たす可能性もあるとの見方が示されています。OpenAI もまた、2027 年の IPO に向けた準備を進めていると伝えられています。
※当記事内のリンクを通じて購入が行われた場合、編集部は少額のコミッションを受け取る場合があります。ただし、これは編集の独立性には一切影響しません。
Lucas への連絡先:lucas.ropek@techcrunch.com
原文を表示
As AI models have become more powerful, the potential for those models to be misused has grown — as has a clamor for safety guardrails that can stop such abuse from happening. AI companies must now walk a delicate tight rope between respecting their enterprise customers’ privacy while also watching usage for possible issues.
Sensing an opportunity to one-up its rival Anthropic, OpenAI just announced a privacy-centric safety approach to monitoring for misuse. The company is previewing a new service to select customers that it calls Private Safety Processing. This is an automated system that watches for potential abuse while simultaneously retaining none of the customer’s data.
This system clearly runs counter to Anthropic’s recently announced data retention policy. The policy, which has aggravated some customers, enables the AI lab to keep user data (all of their sessions — and the conversations therein) for a period of 30 days, when it comes to “covered models.” Those models include all Mythos-class models and “future models with similar capabilities,” the company says.
This policy, which was announced in July, was designed for the purposes of safety allowing the lab to sift and analyze potential impropriety. However, it has deeply concerned some enterprises that handle large amounts of sensitive data and don’t want it harbored (or inspected) by the AI lab.
OpenAI — like most other AI companies — already afford customers a relative level of privacy by adhering to a policy known as Zero Data Retention. ZDR uses agents within the OpenAI API to monitor for abuse on a per session basis. In this way, customer data isn’t retained by the company but companies are still able to scan for bad activity without the need for human intervention. It’s worth noting that Anthropic also largely abides by ZDR — except when it comes to “covered models,” like Fable.
OpenAI says that Private Safety Processing is a new technology that widens ZDR’s scope. It describes it as a form of long-horizon safety monitoring that assesses the inputs and outputs of multiple conversations — not just one. Again, the monitoring is conducted by an agent, which, if triggered, catches interactions and analyzes them across sessions for signs of potential misuse.
The new tech helps OpenAI detect malicious use of AI that takes place over multiple sessions, a spokesperson told TechCrunch. A bad actor — hypothetically someone trying to engineer malware for a cyberattack — may spread out their requests to avoid detection. Private Safety Processing can analyze those multiple conversations for signs of abuse without human review of a user’s conversations.
In the case where the system is triggered, it may send a “narrowly defined signal” to OpenAI that warns of a specific type of activity, the company says. Based on that signal, OpenAI can then decide whether “enforcement is necessary,” it says. If so, OpenAI will reach out to the customer for more context or to work with them on the issue and a customer may choose to share data with OpenAI at their discretion, the spokesperson said.
By contrast, Anthropic notes that human review of customer data can occur, but only “through a controlled access path” that involves “a small set of approved reviewers.” Every one of those review sessions is “recorded in a tamper-proof log that reviewers cannot suppress or modify,” the company says.
The corporate competition between OpenAI and Anthropic is tense at the moment, with both companies looking for any opportunity to gain an advantage on the other. A recent report showed that OpenAI’s Q2 grew more slowly than Anthropic. Anthropic’s annualized revenue run rate is now reportedly $65 billion. Anthropic investors have said it could IPO at $2 trillion, while OpenAI is also working on its IPO.
*When you purchase through links in our articles, we may earn a small commission. This doesn’t affect our editorial independence.*
Lucas is a senior writer at TechCrunch, where he covers artificial intelligence, consumer tech, and startups. He previously covered AI and cybersecurity at Gizmodo.
You can contact Lucas by emailing lucas.ropek@techcrunch.com.
関連記事
今日のまとめ
AIデイリーブリーフで今日の重要ニュースをまとめ読み