OpenAI、開発中のモデル「Astra」がサイバーセキュリティに重大な能力を持つと判断し開発を停止
本文の状態
日本語全文を表示中
詳細モードで約2分の本文を読めます。
同じ出来事の情報源
この情報源を基点に整理
The Verge AI
OpenAI は開発中のモデル「Astra」が極めて強力なサイバーセキュリティ機能を持つ可能性を理由に、その展開を一時的に停止する方針を示した。
Continue in AI NEW LAB
このニュースを、実務の判断につなげる
AI NEW LABで、試したことや先に確認したい条件を共有できます。まずはログインなしで読めます。
AI NEW LABで論点を見るAI深層分析を開く2026年8月8日 05:16
AI深層分析
キーポイント
Astra モデルのリスク評価
OpenAI は現在開発中の「Astra」というモデルが、重要なサイバーセキュリティ機能(critical cybersecurity capabilities)を備えている可能性があると判断している。
展開停止の決定
同社の強力な能力に対する懸念から、Astra モデルの公開や実装に関する進捗を一時的に抑制する「ブレーキ」をかける措置が取られた。
安全対策への優先度転換
技術的な性能向上よりも、潜在的なセキュリティリスクの管理と制御が最優先課題として扱われている現状が示された。
Astraモデルの開発一時停止
OpenAIは新モデル「Astra」について、社内セキュリティ基準を満たしていないため内部活動の一時停止を発表した。
サイバー攻撃能力への懸念
同社によるとAstraはコード作成とサイバーセキュリティにおいて大幅な進歩を遂げており、その強力さが止める要因となった。
重要な引用
OpenAI says its in-development Astra model may have 'critical' cybersecurity capabilities.
"it is pausing 'internal activities' around an in-development AI model, Astra"
"significant advancements in agentic coding and cybersecurity"
Under our Preparedness Framework, a model reaches the Critical cybersecurity threshold if it can identify and develop functional zero-day exploits of all severity levels in many hardened real-world critical systems without human intervention, or can devise and execute end-to-end novel strategies for cyberattacks against hardened targets given only a high level desired goal.
編集コメントを表示
編集コメント
開発中のモデルが「強すぎる」ことを理由に公開を停止する事例は、AI セーフティの重要性が技術革新そのものよりも優先される局面を示唆している。OpenAI の判断は、単なる遅延ではなく、潜在的なリスクに対する慎重な姿勢の表れとして捉えられる。
Source Article
元記事を日本語で読む
本文に関係しない購読案内、埋め込み通知、サイト内プロモーションは除いています。
OpenAI、開発中の「Astra」モデルがサイバーセキュリティにおいて「決定的な」能力を持つ可能性を指摘
OpenAI は現在開発中の「Astra」というモデルについて、サイバーセキュリティ分野で極めて重要な能力を備えている可能性があることを明らかにしました。
著者:Jay Peters
2026 年 8 月 8 日 GMT+9 3:40


画像:The Verge
Jay Peters は、テクノロジーやゲームなどを中心に報道するシニア記者です。2019 年に The Verge に加入する前は Techmeme で約 2 年間勤務していました。
OpenAI は、開発中の AI モデル「Astra」に関する社内活動を一時停止すると発表しました。その理由は、同社が新たに設けたセキュリティ基準を満たしていないためです。
この発表は直近の出来事を受けて行われたものです。OpenAI のモデルが誤って Hugging Face をハッキングしたことが明らかになったばかりでした。また、Anthropic や Meta も、自社の AI モデルが制御不能となり他組織に侵入したことを認めています。
OpenAI によると、Astra と呼ばれる同社モデルの内部評価では、「エージェントによるコーディングとサイバーセキュリティにおいて顕著な進歩」が見られたといいます。同社はさらに、「これらの結果と専門家の評価を踏まえ、昨夜、我々の準備フレームワークの下で重要なサイバー能力を排除できないと判断した」と述べています。
OpenAI が「重要(critical)」なサイバーセキュリティの閾値とみなす基準は以下の通りです。
「準備態勢フレームワーク」に基づき、モデルが以下のいずれかの能力を備えた場合、「サイバーセキュリティの臨界点」に達したと判断されます。
- 人間の介入なしで、多くの堅牢な実世界の重要システムにおいて、あらゆる深刻度のゼロデイ脆弱性を特定し、機能的な攻撃コードを開発できること。
- 高レベルの目標のみを与えられ、堅牢な標的に対するサイバー攻撃に対して、独自のエンドツーエンド戦略を考案・実行できること。
OpenAI は、「Astra は Hugging Face の侵害には関与していない」と述べています。
同社は今回の件を受け、「より高能力なモデルおよび関連する活動については、セキュリティ制御を強化する」と発表しました。また、Astra に関しては「すべてのエージェント型アプリケーションにおいて、リスクのある行動やアライメントのズレに対する包括的な監視」を導入したとしています。
The Verge Daily
最も重要なニュースを毎日お届けする無料ダイジェスト。
Email (必須)
原文を表示
OpenAI says its in-development Astra model may have ‘critical’ cybersecurity capabilities.
OpenAI says its in-development Astra model may have ‘critical’ cybersecurity capabilities.
by Jay Peters
Aug 8, 2026, 3:40 AM GMT+9


Image: The Verge
Jay Peters
is a senior reporter covering technology, gaming, and more. He joined The Verge in 2019 after nearly two years at Techmeme.
OpenAI says it is pausing “internal activities” around an in-development AI model, Astra, because it doesn’t yet meet new security standards the company is putting in place. The announcement follows its recent disclosure that OpenAI models accidentally hacked Hugging Face. Anthropic and Meta have also since admitted that they had AI models that went rogue and breached other organizations.
Recent internal evaluations of an OpenAI model called Astra indicate that it offers “significant advancements in agentic coding and cybersecurity,” according to the company. “These results, in addition to expert assessments, have led us to conclude last night that we cannot rule out critical cyber capabilities under our Preparedness Framework.”
Here is how OpenAI defines a “critical” cybersecurity threshold:
Under our Preparedness Framework, a model reaches the Critical cybersecurity threshold if it can identify and develop functional zero-day exploits of all severity levels in many hardened real-world critical systems without human intervention, or can devise and execute end-to-end novel strategies for cyberattacks against hardened targets given only a high level desired goal.
Astra was “not involved” in the Hugging Face breach, OpenAI says.
OpenAI will implement “stricter security controls for higher-capability models and associated activities,” according to the post. For Astra, it has also implemented “universal monitoring” for “risky actions and misalignment across all agentic applications.”
Follow topics and authors from this story to see more like this in your personalized homepage feed and to receive email updates.
- Jay Peters
-
-
-
-
-
The Verge Daily
A free daily digest of the news that matters most.
Email (required)
関連記事
今日のまとめ
AIデイリーブリーフで今日の重要ニュースをまとめ読み