OpenAI、GPT-5.6 の価格を値下げ
本文の状態
日本語全文を表示中
詳細モードで約7分の本文を読めます。
同じ出来事の情報源
この情報源を基点に整理
TLDR AI
OpenAI は GPT-5.6 Luna と Terra の価格を大幅に引き下げ、Sol モードの処理速度を向上させることで、AI 利用コストの削減と高速化を実現した。
AI深層分析を開く2026年7月31日 23:14
AI深層分析
キーポイント
GPT-5.6 Luna と Terra の価格改定
OpenAI は GPT-5.6 Luna の料金を 80% 引き下げ、Terra の料金を 20% 引き下げたと発表した。
GPT-5.6 Sol の高速化と新モード導入
API に「Fast mode」を導入し、標準処理に比べて最大 2.5 倍の速度で GPT-5.6 Sol を提供するとした。
コスト効率とスケーラビリティの向上
Luna モデルは高品質なまま大量処理を可能にし、企業による大規模 AI アプリケーションの実装を容易にするとしている。
Lunaモデルの性能とコスト効率
Lunaは1年前の最前線クラス相当のパフォーマンスを、タスクあたり約6セントで、ほぼ9倍の速度で提供している。専門作業における評価では、Fable 5を上回り、タスクあたりの推定コストが99%低い結果を示した。
柔軟なワークフロー設計とモデル選択
企業は必要な成果と品質基準を定義し、評価を通じて知能の追加が結果にどう寄与するかを判断して最適なモデルを選定できる。例えばコーディングではSolで計画を立て、Lunaで実装やテストを実行するなど、工程ごとに異なるバランスでモデルを組み合わせて使用可能だ。
重要な引用
GPT‑5.6 Luna, our fastest and most affordable model, will cost 80% less
Fast mode now delivers up to 2.5× faster speeds than Standard processing at twice the price
Making advanced intelligence more abundant and affordable is central to OpenAI's mission
Luna delivers performance comparable to models that were frontier-class a year ago at roughly 6 cents on the dollar per task, and at nearly nine times the speed.
編集コメントを表示
編集コメント
OpenAI はモデルの効率化技術を即座に価格改定と速度向上という形で顧客還元しており、市場競争が激化する中でコストパフォーマンスを武器にした攻めの姿勢を示している。企業利用者はワークフローの要件に合わせてモデルを選択する柔軟性がさらに高まったと言える。
Source Article
元記事を日本語で読む
本文に関係しない購読案内、埋め込み通知、サイト内プロモーションは除いています。
昨日、GPT-5.6 がどのようにして実行効率を高め、自分自身をより軽くしたかについてお伝えしました。今日は、その恩恵を顧客に還元するため、GPT-5.6 Luna と GPT-5.6 Terra の価格を引き下げ、API 上の GPT-5.6 Sol ではパフォーマンスを向上させます。これらのアップデートにより、AI への投資からより多くの価値を得られ、時間が必要な場面でも迅速に対応できるようになります。
本日付で、最速かつ最もコスト効率に優れたモデルである GPT-5.6 Luna の料金は 80% 引きとなり、日常業務に適したバランス型の GPT-5.6 Terra も 20% 値下げされます。Luna と Terra の価格改定は、Codex や ChatGPT Work を利用する際の課金カウントにも反映されます。Luna は、高品質を維持しながら大量の作業を処理する企業にとって、はるかにコスト効率の良い選択肢です。ツールを活用して多段階のワークフローを完結できるため、大規模展開が現実的な AI アプリケーションの範囲も広がります。
高度な知能をより身近で手頃なものにすることは、AGI が人類全体に利益をもたらすことを目指す OpenAI のミッションの中核です。今回の変更は、その約束を実践に移したものです。これは、モデルの構築・提供・活用方法における長年の改善の成果を反映したものです。
API には「Fast モード」も新たに導入しました。これは従来の「Priority Processing」に代わるものです。GPT-5.6 Sol の Fast モードでは、知能レベルはそのままに、標準処理よりも最大 2.5 倍の速度を実現します。料金は 2 倍になりますが、その分高速化されています。Fast モードは後方互換性があり、優先タグ付きのリクエストは自動的にこのモードで処理されます。
成果に応じた適切な知能の選択
AI を効率的に活用するには、まず「何を実現したいか」を明確にする必要があります。エラーのコストや緊急性、規模などによって、求められる知能レベル・速度・信頼性・コストのバランスは異なります。また、ワークフローのステップごとに最適なバランスが変わることもあります。
GPT-5.6 は、こうした最適化の可能性を広げます。例えば Luna モデルは、1 年前に最先端クラスだったモデルと同等のパフォーマンスを、タスクあたりのコストを約 6% に抑えながら提供します。さらに処理速度はほぼ 9 倍です。専門的な業務評価である「Agents' Last Exam」では、Luna は Fable 5 を上回っており、推定されるタスクあたりのコストは約 99% 低いとされています。
実際の運用では、まず必要な成果と品質基準を定義し、評価を通じて「追加の知能が結果に大きく寄与する箇所」と「より高速・低コストな処理で同等の品質が得られる箇所」を見極めることができます。例えばコーディングのワークフローでは、Sol モデルを使って不確実性を解消し計画を立てた後、Luna モデルで具体的な変更を実装し、テストの作成・実行、結果の評価を行うといった使い分けが可能です。また、別のワークフローでは異なるバランスが求められることもあります。
GPT-5.6 ファミリーは、その選択肢の範囲を広げます。企業は、価値が生まれる各段階で最大限に有用な知能を適用しながら、その価値に見合った適切な価格を支払うことができます。
この柔軟性を実現する第一歩は、モデル背後にあるすべての層をより効率的にすることです。
効率のフロンティアをどう押し広げるか
当社の効率性の優位性は、モデル自体の改善、それらを実行する推論システム、そしてツールや文脈と接続するエージェント・ハネス(harness)の強化によって生まれています。GPT-5.6 モデルは、作業においてより直接的な経路をたどります。
適切なルーティングによりハードウェアの稼働率が高まり、最適化された生成ソフトウェアがトークンをより効率的に出力します。また、文脈管理の知能化によって、エージェントは完了した作業の繰り返しを避けることができます。これらの改善を組み合わせることで、同じ計算資源でより多くの有用な作業を完遂できるようになり、結果を得るために必要な時間、トークン数、コストを削減できます。
GPT-5.6 の「Sol」は、次の収益向上の機会を特定し実現する上で、ますます重要な役割を果たしています。人間の主導プロセスの中で、Sol は生産用カーネルの書き換えと最適化を自律的に行い、トークン生成の改善のために数百回の実験を設計・実行しました。また、トレーニングの監視も担当し、問題が発生した際には介入します。このカーネル関連の取り組みにより、モデル提供にかかるエンドツーエンドのコストが 20% 削減され、実験の結果、トークン生成の効率は 15% 以上向上しました。こうした活動は継続されており、より緊密なフィードバックループを形成しています。モデルが高度化し自律性が向上するほど、効率改善のスピードも加速します。
GPT-5.6 の背後にあるエンジニアリングの詳細については、こちらをご覧ください。
スケーラビリティに特化した計算戦略
膨大な知能への需要に応えるには、計算リソースの増強と、より生産的な計算環境の両方が必要です。私たちは堅牢なインフラポートフォリオを構築し、各ワークロードに最適なシステムを割り当てています。このアプローチは、価格性能曲線の両端に対応しています。低コスト側では、新たに設定された Luna と Terra の料金プランにより、大規模な高ボリューム処理が経済的に可能になりました。一方、最前線(フロンティア)向けには、レスポンス時間が重要な場合に Sol への高速アクセスを提供する「Fast モード」を API ユーザーに提供しています。
企業は、最も重要な業務のスピードを犠牲にすることなく、AI を日常業務により多く取り入れることができます。大規模な文書分析や顧客対応の分類、そして日常的な実装作業も、広範囲で実行するコストが抑えられます。一方、複雑な Sol ワークロードについては、プレミアム料金が正当化される場合に処理速度を向上させることが可能です。
これらの効果は複合的に積み上がります。人間の主導によるプロセスにおいて、より高性能なモデルが技術チームの支援となり、次世代の改善点を早期に見つけることで、性能向上とコスト削減への道筋を短縮します。当社の戦略は引き続き、能力と効率性の両面での進化に焦点を当てており、各世代の知能システムが、より低いコストでより多くの業務を処理できるようになります。
利用開始と料金
GPT-5.6 Terra と Luna は、ChatGPT Work、Codex、および OpenAI API で引き続きご利用いただけます。ChatGPT Work と Codex では、Free および Go プランのユーザーが Terra にアクセスでき、Plus、Pro、Business、Enterprise プランのユーザーは Terra と Luna の両方を選択可能です。
7 月 30 日より、API 料金は Terra が入力トークン 100 万あたり 2 ドル、出力トークン 100 万あたり 12 ドルとなります。Luna はそれぞれ 0.20 ドルと 1.20 ドルです。Sol の料金は変更されません。ChatGPT と Codex のサブスクリプション料金およびクォータ予算も据え置きますが、Terra と Luna の利用で消費されるクレジットの量が減少します。AWS における料金改定は本日順次適用されます。
GPT-5.6 Sol の「高速モード」が、API における優先処理に代わり、Codex の /fast と整合するようになりました。すでに「優先」としてタグ付けされた API リクエストは引き続き動作します。
原文を表示
Yesterday, we shared how GPT‑5.6 helped make itself more efficient to run. Today, we’re passing those gains on to customers with lower prices for GPT‑5.6 Luna(opens in a new window) and Terra(opens in a new window) and faster performance with GPT‑5.6 Sol in the API. Together, these updates help customers get more from every dollar they invest in AI and move faster when time matters.
Starting today, GPT‑5.6 Luna, our fastest and most affordable model, will cost 80% less, while GPT‑5.6 Terra, our balanced model for everyday work, will cost 20% less. These lower prices for Luna and Terra are also reflected in how usage is counted against paid subscriptions when using Codex and ChatGPT Work. Luna gives businesses a far more cost-effective way to handle high-volume work at very high levels of quality. It can use tools and complete multi-step workflows, making a broader range of AI applications practical to run at scale.
Making advanced intelligence more abundant and affordable is central to OpenAI’s mission to ensure AGI benefits all of humanity. These changes put that commitment into practice. They reflect years of improvements in how our models are built, served, and put to work.
We’re also introducing Fast mode in the API, which replaces our Priority Processing offering. For GPT‑5.6 Sol, Fast mode now delivers up to 2.5× faster speeds than Standard processing at twice the price, with no change in intelligence. Fast mode is backward compatible: requests tagged priority will automatically use Fast mode.
1 of 6
Matching intelligence to the outcome
Using AI efficiently begins with the outcome. The stakes, cost of error, urgency, and scale determine the right balance of intelligence, speed, reliability, and cost. That balance can change from one step of a workflow to the next.
GPT‑5.6 gives businesses much more room to optimize that equation. Luna delivers performance comparable to models that were frontier-class a year ago at roughly 6 cents on the dollar per task, and at nearly nine times the speed. On professional work, as measured by Agents’ Last Exam, Luna outperforms Fable 5 at an estimated cost per task nearly 99% lower.
In practice, businesses can define the outcome and quality standard they need, then use evaluations to determine where additional intelligence materially improves the result and where faster, lower-cost processing can deliver the same quality. A coding workflow, for example, might use Sol to resolve uncertainty and define the plan, then use Luna to implement well-specified changes, write and run tests, and evaluate the results. Another workflow may call for a different balance.
The GPT‑5.6 family expands the range of those choices. Businesses can apply the maximum useful intelligence at every stage while paying the right price for the value it creates.
Delivering that flexibility starts with making every layer behind the models more efficient.
How we advance the efficiency frontier
Our efficiency edge comes from improving the models, the inference systems that run them, and the agentic harness that connects them to tools and context. GPT‑5.6 models take a more direct path through work. Better routing keeps hardware productive, optimized production software generates tokens more efficiently, and smarter context management helps agents avoid repeating completed work. Together, these improvements let us complete more useful work with the same compute, reducing the time, tokens, and cost required for each result.
GPT‑5.6 Sol is increasingly helping us find and deliver the next round of gains. Within a human-led process, Sol autonomously rewrote and optimized production kernels, designed and ran hundreds of experiments to improve token generation, and monitored training, intervening when problems arose. The kernel work helped reduce the end-to-end cost of serving the model by 20%, while its experiments increased token-generation efficiency by more than 15%. This work continues, creating a tighter feedback loop: as our models improve and are able to work more autonomously, our ability to improve efficiencies accelerates. Read more about the engineering behind GPT‑5.6.
A compute strategy built for scale
Meeting demand for abundant intelligence requires both more compute and more productive compute. We are building a resilient infrastructure portfolio and matching each workload to the systems best suited to run it. That approach supports both ends of the price-performance curve. At the lower-cost end, the new Luna and Terra prices make high-volume work economical at much greater scale. At the frontier end, Fast mode gives API customers faster access to Sol when response time is important.
Enterprises can move more AI into everyday operations without sacrificing speed on their most consequential work. Large-scale document analysis, customer-interaction classification, and routine implementation can become economical to run broadly, while complex Sol workloads can move faster when the premium is justified.
The gains can compound. Within a human-led process, more capable models help our technical team find the next generation of improvements, shortening the path to better performance and lower costs. Our strategy remains focused on advancing both capability and efficiency so each generation of intelligence can accomplish more work at a lower cost.
Availability and pricing
GPT‑5.6 Terra and Luna remain available in ChatGPT Work, Codex, and the OpenAI API. In ChatGPT Work and Codex, Free and Go users can access Terra, while Plus, Pro, Business, and Enterprise users can choose Terra and Luna.
Starting July 30, API pricing is $2 per million input tokens and $12 per million output tokens for Terra, and $0.20 per million input tokens and $1.20 per million output tokens for Luna. Sol pricing remains unchanged. ChatGPT and Codex subscription prices and quota budgets remain unchanged, while Terra and Luna usage now consumes fewer credits. Pricing changes will begin rolling out in AWS later today.
Fast mode for GPT‑5.6 Sol replaces Priority Processing in the API and aligns with /fast in Codex. Existing API requests tagged priority will continue to work. View complete API pricing details.
関連記事
News to Guide
ニュースの次に確認する
発表内容を、現在の料金や仕様と照らし合わせられる関連ガイドです。
今日のまとめ
AIデイリーブリーフで今日の重要ニュースをまとめ読み