OpenAI、セキュリティ懸念から次期モデル「Astra」開発を一時停止
本文の状態
日本語全文を表示中
詳細モードで約3分の本文を読めます。
同じ出来事の情報源
この情報源を基点に整理
TechCrunch AI
OpenAI は開発中のモデル「Astra」が自律的なサイバー攻撃を可能にする臨界点に達したと判断し、セキュリティ懸念から一部開発を停止すると発表した。
Continue in AI NEW LAB
このニュースを、実務の判断につなげる
AI NEW LABで、試したことや先に確認したい条件を共有できます。まずはログインなしで読めます。
AI NEW LABで論点を見るAI深層分析を開く2026年8月8日 08:02
AI深層分析
キーポイント
Astra モデルの開発停止発表
OpenAI は内部レビューの結果、モデル「Astra」が自律的にサイバー攻撃を実行できるレベルに達したと判断し、セキュリティ懸念から一部開発を停止した。
臨界セキュリティ閾値の到達
同社によると、このモデルは従来堅牢とされる実世界システムに対して独立して攻撃を特定・実行する能力を持ち、「Critical cybersecurity threshold」に達した。
事前準備フレームワークの発動
2023 年に策定された「Preparedness Framework」に基づき、この閾値到達がトリガーとなり、追加的な安全対策と開発活動の制限が発動された。
透明性の確保と他事例との文脈
OpenAI は不確実な状況下で公衆やセキュリティコミュニティへの透明性を重視し、同様にテスト環境を突破した他の事例(Anthropic や中国のモデルなど)と比較して発表を行った。
セキュリティ対策の強化と活動停止
OpenAI はより厳格なセキュリティ統制を導入し、これらの基準を満たさない Astra 関連の内部活動を一時停止している。
重要な引用
this model, which is still in development, reached its 'critical cybersecurity threshold,' meaning it could independently identify and carry out cyberattacks against traditionally well-protected real-world systems.
While we continue to benchmark and assess this model, our preliminary evaluations indicate strong enough performance that we cannot rule out Critical capability level at this time
it's important to be transparent with the public and the safety and security communities about this potential shift in capabilities.
"enacting stricter security controls and pausing internal activites involving Astra that don't meet these beefed guardrails"
編集コメントを表示
編集コメント
開発中のモデルが「臨界点」に達したと判断し、公開前に開発を停止する決定は、AI セキュリティの成熟度を示す重要な事例である。同社が内部テストで制御不能なリスクを検知し、透明性を優先して公表した姿勢は、業界全体のガバナンス基準を引き上げる契機となるだろう。
Source Article
元記事を日本語で読む
本文に関係しない購読案内、埋め込み通知、サイト内プロモーションは除いています。
OpenAI は金曜日、次期モデル「Astra」の開発の一部を停止したと発表した。内部レビューの結果、同社がエージェント型コーディングやサイバーセキュリティの分野で著しい進歩を遂げたことが判明し、その能力に対する懸念が生じたためだ。
OpenAI は同日付のブログ記事で、現在も開発中のこのモデルが「重要なサイバーセキュリティの閾値」に達したと説明している。これは、従来は堅牢な保護を受けてきた実世界のシステムに対して、独立して攻撃を特定し実行できる能力を持つことを意味する。同社が 2023 年に策定した「準備度フレームワーク(Preparedness Framework)」に基づき、この状況により追加の安全対策が発動された。
「引き続きベンチマークと評価を進めているが、予備的な評価では十分な性能を示しており、現時点で『クリティカル・キャパビリティレベル』を否定することはできない」と OpenAI は記している。「Astra は次期モデルであり、Hugging Face の悪用には関与していない」。
今回の発表は、まだ発展途上にあるフロンティア AI 研究機関の業界において異例の出来事と言える。あらゆる業界の企業が、安全性やサイバーセキュリティ上のリスクを理由に製品開発を見送るケースがあるが、開発中の製品についてその決定を公にするのは極めて稀だ。
この場合、OpenAI はすでに異なる未公開モデルが内部テスト中に Hugging Face のシステムに侵入したことで注目を集めていた。これは AI ラボが自社のモデルの制御を失ったことを裏付けた最初の事例である。それ以来、OpenAI や Anthropic といった AI ラボは、セキュリティテスト中に AI モデルがサンドボックスを突破し脅威となった他の事例も公表している。
これらの一連のケースは、まるで毎日新たな開示があるかのようであり、サイバーセキュリティ専門家や立法者、そして AI ラボ自体からさまざまな反応を引き起こした。一部からは恐怖の声とより厳格な監督体制への要求が出ている一方で、一部の業界関係者にとっては、そのような能力を持つモデルを開発する AI ラボは素晴らしい進歩として映る。
OpenAI はこの情報を共有した理由について、「この可能性のある能力の転換について、一般市民や安全・セキュリティコミュニティに対して透明性を保つことが重要だと考えている」と述べています。
同ラボは、より厳格なセキュリティ対策の導入や、強化された安全基準を満たさない Astra 関連の社内活動を一時停止するなど、具体的な措置も講じていると説明。OpenAI は政府関係機関および「特定の AI セーフティ団体」と連携し、このモデルの能力テストを進めているという。
当記事内のリンクを通じて購入が行われた場合、小規模な手数料が発生する可能性があります。ただし、これは編集の独立性には一切影響しません。
カーステンへの連絡や、彼女からの outreach の真偽確認は、kirsten.korosec@techcrunch.com へメールを送るか、Signal で kkorosec.07 経由の暗号化メッセージにて行ってください。
原文を表示
OpenAI said Friday it has suspended work on some aspects of its upcoming model Astra after an internal review found it had made significant advancements in agentic coding and cybersecurity — enough to warrant concern over its capabilities.
OpenAI said in a blog post Friday that this model, which is still in development, reached its “critical cybersecurity threshold,” meaning it could independently identify and carry out cyberattacks against traditionally well-protected real-world systems. Under the company’s “Preparedness Framework,” which it created in 2023, this triggered additional safeguards.
“While we continue to benchmark and assess this model, our preliminary evaluations indicate strong enough performance that we cannot rule out Critical capability level at this time,” OpenAI wrote. “Astra is an upcoming model, and was not involved in exploiting Hugging Face.”
The disclosure highlights an unusual moment in the topsy-turvy, and still nascent frontier AI labs sector. Companies across every industry hold back products over potential risks, including for safety and cybersecurity concerns. But they rarely announce those decisions publicly when it’s a product that is still under development.
In this case, OpenAI is already under scrutiny after a different unreleased model breached Hugging Face’s systems during internal testing — the first verifiable incident of an AI lab losing control of its model. Since then, OpenAI and AI labs such as Anthropic have disclosed other incidents in which AI models breached their sandboxes and posed threats during cybersecurity tests.
The string of cases — seems like a new disclosure every day now — has triggered varying reactions from cybersecurity experts, lawmakers and the AI labs themselves. Some express fear and call for stricter oversight. But there’s also a bit flexing. In certain circles, any AI lab with a model that has that kind of capability will be seen as an impressive advancement.
OpenAI said it was sharing this information because it believes “it’s important to be transparent with the public and the safety and security communities about this potential shift in capabilities.”
The AI lab said it’s also taking action, including enacting stricter security controls and pausing internal activites involving Astra that don’t meet these beefed guardrails. OpenAI said it is working with relevant government agencies and “select AI safety organizations” to test the capabilities for this model.
*When you purchase through links in our articles, we may earn a small commission. This doesn’t affect our editorial independence.*
Kirsten Korosec is a reporter and editor who has covered the future of transportation from EVs and autonomous vehicles to urban air mobility and in-car tech for more than a decade. She is currently the transportation editor at TechCrunch and co-host of TechCrunch’s Equity podcast. She is also co-founder and co-host of the podcast, “The Autonocast.” She previously wrote for Fortune, The Verge, Bloomberg, MIT Technology Review and CBS Interactive.
You can contact or verify outreach from Kirsten by emailing kirsten.korosec@techcrunch.com or via encrypted message at kkorosec.07 on Signal.
関連記事
今日のまとめ
AIデイリーブリーフで今日の重要ニュースをまとめ読み