OpenAI、GPT-5.6 で価格性能比の限界を前進と発表
OpenAI は GPT-5.6 の効率化成果を反映し、Luna と Terra モデルの価格を大幅に引き下げるとともに Sol モデルの高速モードを導入して、顧客が AI 投資からより多くの価値を得られるようにした。
AI深層分析を開く2026年7月31日 02:51
AI深層分析
キーポイント
GPT-5.6 Luna の価格大幅削減
OpenAI は最も高速で安価なモデルである GPT-5.6 Luna の料金を 80% 引き下げ、高品質かつ大量のワークフロー処理を低コストで行える環境を整えた。
GPT-5.6 Terra の価格引き下げ
日常的な業務に適したバランス型モデルである GPT-5.6 Terra の料金を 20% 引き下げ、Codex や ChatGPT Work の利用料金体系にも反映させた。
API における高速モードの導入
OpenAI は API に「Fast mode」を導入し、標準処理より最大 2.5 倍速く動作する GPT-5.6 Sol の提供を開始した。
AGI 実現へのコミットメント
高度な知能をより多くの人々が利用可能にすることを目指し、モデルの構築・運用技術の向上が価格低下と性能向上につながったと説明している。
Lunaモデルの性能とコスト効率
Lunaは1年前の最前級モデルに匹敵するパフォーマンスを提供し、タスクあたりのコストを約6セント(1/100)に抑えつつ、処理速度を9倍向上させた。
重要な引用
Luna gives businesses a far more cost-effective way to handle high-volume work at very high levels of quality.
Making advanced intelligence more abundant and affordable is central to OpenAI's mission to ensure AGI benefits all of humanity.
Luna delivers performance comparable to models that were frontier-class a year ago at roughly 6 cents on the dollar per task, and at nearly nine times the speed.
GPT‑5.6 Sol is increasingly helping us find and deliver the next round of gains.
編集コメントを表示
編集コメント
今回の発表は、モデルの技術的効率化が即座に顧客への価格還元として現れた好例である。特に Luna モデルの 80% の値下げは、大規模な AI アプリケーション展開におけるコスト構造を根本から変える可能性を秘めている。
Source Article
元記事を日本語で読む
本文に関係しない購読案内、埋め込み通知、サイト内プロモーションは除いています。
昨日、GPT-5.6 が自身の実行効率を高める方法についてお伝えしました。今日は、その成果をお客様に還元します。GPT-5.6 Luna と GPT-5.6 Terra の利用料金を引き下げるとともに、API 上の GPT-5.6 Sol ではパフォーマンスを向上させました。
これらのアップデートにより、お客様は AI への投資からより多くの価値を得られ、時間が必要な場面でもスピードを上げることができます。
本日付で、最速かつ最もコストパフォーマンスに優れたモデルである GPT-5.6 Luna の料金が 80% 引きになります。また、日常業務に適したバランス型のモデル GPT-5.6 Terra も 20% 値下げされます。Luna と Terra の価格改定は、Codex や ChatGPT Work を利用する際の課金カウントにも反映されます。
Luna は、高品質を維持しながら大量の処理をこなすための、よりコスト効率の良い手段を提供します。ツールを活用し、多段階のワークフローを完遂できるため、大規模展開が現実的な AI アプリケーションの範囲を広げます。
高度な知能をより豊富に、かつ手頃な価格で提供することは、AGI が人類全体に恩恵をもたらすことを目指す OpenAI のミッションの中核です。今回の変更は、その約束を実践するものです。これは、モデルの構築・提供・活用における長年の改善の成果を反映したものです。
API には「Fast モード」も新たに導入し、従来の「優先処理」オプションを置き換えます。GPT-5.6 Sol の Fast モードでは、知能レベルは維持したまま、標準処理より最大 2.5 倍の速度で動作します。料金は約 2 倍ですが、その分圧倒的なスピードが得られます。Fast モードは後方互換性があり、優先タグ付きのリクエストは自動的に Fast モードとして処理されます。
成果に合わせた知能の選択
AI を効率的に活用するには、まず「目指す成果」を明確にすることが出発点です。リスクの大きさ、エラーのコスト、緊急性、そして規模に応じて、知能レベル・速度・信頼性・コストのバランスは変わります。このバランスは、ワークフローのステップごとに最適化できるものです。
GPT-5.6 は、こうした最適化の余地を大幅に広げます。例えば Luna モデルは、1 年前なら最先端クラスだったモデルと同等のパフォーマンスを、タスクあたり約 6 セント(従来比の約 1/16)で提供し、速度も約 9 倍です。また、Agents' Last Exam で測定される専門業務においては、Fable 5 を上回る性能を発揮しながら、タスクあたりの推定コストは約 99% 削減できます。
実際の運用では、企業がまず必要な成果と品質基準を定義し、評価を通じて「追加の知能が結果に大きく寄与する場面」と「より高速・低コストな処理で同等の品質が得られる場面」を見極めます。例えばコーディングワークフローなら、Sol モデルを使って不確実性を解消し計画を立てた後、Luna モデルで具体的な変更を実装し、テストの作成・実行、結果の評価を行います。別のワークフローでは、また異なるバランスが必要になるでしょう。
GPT-5.6 シリーズは、これらの選択肢の範囲をさらに広げます。企業は、価値が生まれる各段階で最大限に有用な知能を活用しつつ、その価値に見合った適切なコストで利用できるようになります。
この柔軟性を実現する鍵は、モデル背後にあるすべてのレイヤーをより効率的にすることにあります。
効率性のフロンティアをどう切り拓くか
私たちの効率化の優位性は、モデル自体の改善、それらを実行する推論システム、そしてツールや文脈と接続するエージェント・ハネス(harness)の強化によって得られています。GPT-5.6 モデルは、作業に対してより直接的なアプローチを採用します。
適切なルーティングによりハードウェアの稼働率を高め、最適化された生成ソフトウェアがトークンをより効率的に出力し、賢明なコンテキスト管理によってエージェントが完了した作業の繰り返しを防ぎます。これらの改善を組み合わせることで、同じ計算リソースでより多くの有用な作業を完遂できるようになり、結果を得るために必要な時間、トークン数、およびコストを削減できます。
GPT-5.6 の Sol は、次の段階の性能向上を見出し、実現する上でますます重要な役割を果たしています。人間の主導プロセスの中で、Sol は生産用カーネルの書き換えと最適化を自律的に行い、トークン生成の改善のために数百回の実験を設計・実行しました。また、トレーニング中の監視も担当し、問題が発生した際には介入します。このカーネルに関する作業により、モデル提供にかかるエンドツーエンドのコストが 20% 削減され、実験の結果、トークン生成効率は 15% 以上向上しました。こうした取り組みは継続しており、より一層緊密なフィードバックループを構築しています。モデルの性能が向上し、自律的な動作が可能になるほど、効率化を進めるスピードも加速します。
GPT-5.6 の背後にあるエンジニアリングについて詳しく読む
スケーラビリティを見据えた計算戦略
膨大な知能への需要に応えるには、計算リソースの増強と、より生産性の高い計算の実現の両方が必要です。私たちは回復力のあるインフラポートフォリオを構築し、各ワークロードに最も適したシステムを割り当てています。このアプローチは、価格性能曲線の両端を支えています。低コスト側では、新しい Luna と Terra の価格設定により、大規模な高ボリューム処理が経済的に可能になりました。一方、最前線側では、レスポンス時間が重要な場合に API ユーザーが Sol にすばやくアクセスできるよう、「Fast モード」を提供しています。
企業は、最も重要な業務のスピードを犠牲にすることなく、AI を日常業務により多く取り入れることができます。大規模な文書分析や顧客インタラクションの分類、そして日常的な実装作業も、広範囲で実行できるよう経済的になります。一方、複雑な Sol ワークロードについては、プレミアムが正当化される場合に処理速度を向上させることが可能です。
これらの効果は複合的に積み上がっていきます。人間の主導するプロセスにおいて、より高性能なモデルが技術チームの次世代改善策を見つける手助けをし、パフォーマンス向上とコスト削減への道筋を短縮します。当社の戦略は、能力と効率性の両方を高めることに焦点を当て続けており、各世代の知能システムが、より低いコストでより多くの業務を遂行できるようにしています。
利用可能状況と料金
GPT‑5.6 Terra と Luna は、ChatGPT Work、Codex、および OpenAI API で引き続き利用可能です。ChatGPT Work と Codex では、Free および Go ユーザーが Terra にアクセスでき、Plus、Pro、Business、Enterprise ユーザーは Terra と Luna の両方を選択できます。
7 月 30 日より、API 料金は Terra が入力トークン 100 万あたり 2 ドル、出力トークン 100 万あたり 12 ドルとなります。Luna は入力トークン 100 万あたり 0.20 ドル、出力トークン 100 万あたり 1.20 ドルです。Sol の料金は変更されません。ChatGPT と Codex のサブスクリプション料金およびクォータ予算も従来通りですが、Terra と Luna の利用にはより少ないクレジットで済むようになります。AWS での料金改定は本日順次適用されます。
GPT-5.6 Sol の高速モードは、API における優先処理に代わり、Codex の「/fast」と整合性が取れるようになりました。すでに「priority」タグ付きで送信された既存の API リクエストも引き続き動作します。
原文を表示
Yesterday, we shared how GPT‑5.6 helped make itself more efficient to run. Today, we’re passing those gains on to customers with lower prices for GPT‑5.6 Luna(opens in a new window) and Terra(opens in a new window) and faster performance with GPT‑5.6 Sol in the API. Together, these updates help customers get more from every dollar they invest in AI and move faster when time matters.
Starting today, GPT‑5.6 Luna, our fastest and most affordable model, will cost 80% less, while GPT‑5.6 Terra, our balanced model for everyday work, will cost 20% less. These lower prices for Luna and Terra are also reflected in how usage is counted against paid subscriptions when using Codex and ChatGPT Work. Luna gives businesses a far more cost-effective way to handle high-volume work at very high levels of quality. It can use tools and complete multi-step workflows, making a broader range of AI applications practical to run at scale.
Making advanced intelligence more abundant and affordable is central to OpenAI’s mission to ensure AGI benefits all of humanity. These changes put that commitment into practice. They reflect years of improvements in how our models are built, served, and put to work.
We’re also introducing Fast mode in the API, which replaces our Priority Processing offering. For GPT‑5.6 Sol, Fast mode now delivers up to 2.5× faster speeds than Standard processing at twice the price, with no change in intelligence. Fast mode is backward compatible: requests tagged priority will automatically use Fast mode.
1 of 6
Matching intelligence to the outcome
Using AI efficiently begins with the outcome. The stakes, cost of error, urgency, and scale determine the right balance of intelligence, speed, reliability, and cost. That balance can change from one step of a workflow to the next.
GPT‑5.6 gives businesses much more room to optimize that equation. Luna delivers performance comparable to models that were frontier-class a year ago at roughly 6 cents on the dollar per task, and at nearly nine times the speed. On professional work, as measured by Agents’ Last Exam, Luna outperforms Fable 5 at an estimated cost per task nearly 99% lower.
In practice, businesses can define the outcome and quality standard they need, then use evaluations to determine where additional intelligence materially improves the result and where faster, lower-cost processing can deliver the same quality. A coding workflow, for example, might use Sol to resolve uncertainty and define the plan, then use Luna to implement well-specified changes, write and run tests, and evaluate the results. Another workflow may call for a different balance.
The GPT‑5.6 family expands the range of those choices. Businesses can apply the maximum useful intelligence at every stage while paying the right price for the value it creates.
Delivering that flexibility starts with making every layer behind the models more efficient.
How we advance the efficiency frontier
Our efficiency edge comes from improving the models, the inference systems that run them, and the agentic harness that connects them to tools and context. GPT‑5.6 models take a more direct path through work. Better routing keeps hardware productive, optimized production software generates tokens more efficiently, and smarter context management helps agents avoid repeating completed work. Together, these improvements let us complete more useful work with the same compute, reducing the time, tokens, and cost required for each result.
GPT‑5.6 Sol is increasingly helping us find and deliver the next round of gains. Within a human-led process, Sol autonomously rewrote and optimized production kernels, designed and ran hundreds of experiments to improve token generation, and monitored training, intervening when problems arose. The kernel work helped reduce the end-to-end cost of serving the model by 20%, while its experiments increased token-generation efficiency by more than 15%. This work continues, creating a tighter feedback loop: as our models improve and are able to work more autonomously, our ability to improve efficiencies accelerates. Read more about the engineering behind GPT‑5.6.
A compute strategy built for scale
Meeting demand for abundant intelligence requires both more compute and more productive compute. We are building a resilient infrastructure portfolio and matching each workload to the systems best suited to run it. That approach supports both ends of the price-performance curve. At the lower-cost end, the new Luna and Terra prices make high-volume work economical at much greater scale. At the frontier end, Fast mode gives API customers faster access to Sol when response time is important.
Enterprises can move more AI into everyday operations without sacrificing speed on their most consequential work. Large-scale document analysis, customer-interaction classification, and routine implementation can become economical to run broadly, while complex Sol workloads can move faster when the premium is justified.
The gains can compound. Within a human-led process, more capable models help our technical team find the next generation of improvements, shortening the path to better performance and lower costs. Our strategy remains focused on advancing both capability and efficiency so each generation of intelligence can accomplish more work at a lower cost.
Availability and pricing
GPT‑5.6 Terra and Luna remain available in ChatGPT Work, Codex, and the OpenAI API. In ChatGPT Work and Codex, Free and Go users can access Terra, while Plus, Pro, Business, and Enterprise users can choose Terra and Luna.
Starting July 30, API pricing is $2 per million input tokens and $12 per million output tokens for Terra, and $0.20 per million input tokens and $1.20 per million output tokens for Luna. Sol pricing remains unchanged. ChatGPT and Codex subscription prices and quota budgets remain unchanged, while Terra and Luna usage now consumes fewer credits. Pricing changes will begin rolling out in AWS later today.
Fast mode for GPT‑5.6 Sol replaces Priority Processing in the API and aligns with /fast in Codex. Existing API requests tagged priority will continue to work. View complete API pricing details.
AI算出
主要ニュースainew評価高い
AI モデルの性能向上と価格改定という核心的なニュースであり、具体的なバージョン番号と数値データが含まれているため新規性と検索機会が高い。ただし、日本固有の価格や規制に関する言及がないため、日本の関連性は低い。
6つの評価軸を見る
- AI関連度
- 100
- 情報源の信頼性
- 100
- 新規性
- 75
- 調べる価値
- 100
- 重複の少なさ
- 100
- 日本での有用性
- 25
関連記事
News to Guide
ニュースの次に確認する
発表内容を、現在の料金や仕様と照らし合わせられる関連ガイドです。
今日のまとめ
AIデイリーブリーフで今日の重要ニュースをまとめ読み