Google、Gemini 3.6 Flash と 3.5 Flash-Lite を発表しタスク処理時間を半減
本文の状態
日本語全文を表示中
詳細モードで約4分の本文を読めます。
同じ出来事の情報源
この情報源を基点に整理
Artificial Analysis
Google DeepMind は Gemini 3.6 Flash と Gemini 3.5 Flash-Lite をリリースし、タスクあたりの処理時間を半減させつつ、コスト効率を向上させた。
AI深層分析を開く2026年8月1日 14:02
AI深層分析
キーポイント
Gemini 3.6 Flash の性能と速度の両立
知能指数は前モデルと同水準(50点)を維持しつつ、タスクあたりの平均処理時間が 1.3 分(約半減)に短縮され、出力トークン速度が秒間 304 トークンに達した。
Gemini 3.5 Flash-Lite の知能指数向上
軽量モデルである Gemini 3.5 Flash-Lite は、知能指数で 11 ポイント向上し、コスト効率と性能のバランスが改善された。
コスト削減と価格改定の実現
Gemini 3.6 Flash のタスクあたりのコストは約 18% 低下し、入力・出力トークンあたりの単価も引き下げられた。
推論能力の大幅な向上
Gemini 3.5 Flash-Lite は Artificial Analysis Intelligence Index で Gemini 3.1 Flash-Lite よりも 11 ポイント上昇し、特にエージェント評価において顕著な改善が見られる。
タスクあたりの処理時間が半減
トークン効率の向上と出力速度の速さにより、Gemini 3.5 Flash-Lite の平均タスク時間は約 0.6 分となり、前モデルと比較してほぼ半分になった。
重要な引用
Both halve time per task relative to their predecessors and increase token efficiency
Gemini 3.6 Flash records an average time per task of 1.3 minutes, a more than 50% reduction compared to Gemini 3.5 Flash
Gemini 3.5 Flash-Lite records an average time per task of 0.6 minutes, nearly a 50% reduction compared to Gemini 3.1 Flash-Lite (1.0).
Gemini 3.5 Flash-Lite scores 36 on the Artificial Analysis Intelligence Index, up 11 points from Gemini 3.1 Flash-Lite (25).
編集コメントを表示
編集コメント
Gemini シリーズの Flash ラインナップが、速度とコストという実用性の観点で明確な進歩を遂げた。特に知能指数を維持したまま処理時間を半減させた点は、大規模展開における採用判断材料として極めて価値が高い。
Source Article
元記事を日本語で読む
本文に関係しない購読案内、埋め込み通知、サイト内プロモーションは除いています。
Google は「Gemini 3.6 Flash」と「Gemini 3.5 Flash-Lite」の 2 つの新モデルをリリースしました。両モデルとも前世代に比べてタスクあたりの所要時間が半分になり、トークン効率も向上しています。特に Gemini 3.5 Flash-Lite は知能指数が 11 ポイント改善しましたが、Gemini 3.6 Flash の知能レベルは Gemini 3.5 Flash と同等で、大幅な向上はありません。
Google DeepMind が Gemini モデルファミリーの最新アップデートとして発表した 2 つの新モデルについて、リリース前に「知能」「タスクあたりの時間」「コスト」の観点からベンチマークを行いました。

Gemini 3.6 Flash の主な特徴(高推論性能)
➤ 知能指数は Gemini 3.5 Flash と同等: Artificial Analysis の知能指数で Gemini 3.6 Flash は 50 を記録し、Gemini 3.5 Flash と同じスコアです。直近で発表された Muse Spark 1.1 (xhigh, 51) や GPT-5.6 Luna (max, 51) よりわずかに下回ります。Gemini 3.5 Flash と比較すると、指数全体で同程度のスコアを維持しつつ、GDPval-AA v2 では 1421 (+72 ポイント) と改善が見られる一方、HLE では 38% (-3 ポイント) とやや後退しています。
➤ タスクあたりの時間が半分: Gemini 3.6 Flash のタスク平均所要時間は 1.3 分です。これは Gemini 3.5 Flash (2.7 分) に比べて 50% 以上短縮された結果です。トークン効率の向上と出力速度の加速が要因で、リリース前のテストでは 1 秒あたり 304 トークンの出力速度を記録しました。
➤ タスクあたりのコストがわずかに低下: Gemini 3.6 Flash のタスクあたりのコストは約 18% 減少し、0.59 ドルから 0.50 ドルに下がりました。これは出力トークンの使用量が減ったことと、新しい価格設定(入力・出力ともに 100 万トークンあたり 1.50 ドル/7.50 ドル)によるものです。以前の Gemini 3.5 Flash では 1.50 ドル/9.00 ドルでした。

Gemini 3.5 Flash-Lite(高推論型)の主なポイント:
➤ Gemini 3.1 Flash Lite より大幅な知能向上: Gemini 3.5 Flash-Lite は Artificial Analysis のインテリジェンス指数で 36 を記録し、Gemini 3.1 Flash Lite(25)から 11 ポイント向上しました。このスコアは Nemotron 3 Ultra(38)や DeepSeek V4 Flash(最大値 40)には及びませんが、Mistral Medium 3.5(30)を上回ります。Gemini 3.1 Flash Lite と比較した知能の最大の向上はエージェント評価において見られ、GDPval-AA v2 で 1140(+498)、TerminalBench v2.1 で 53.6(+22.5 ポイント)、Tau3-Banking で 16.5%(+7.8 ポイント)と大幅な改善が確認されました。
➤ タスクあたりの時間がほぼ半分: Gemini 3.5 Flash-Lite のタスクあたりの平均時間は 0.6 分で、Gemini 3.1 Flash Lite(1.0 分)と比較して約 50% の短縮を実現しました。これはトークン効率の向上と高速な出力速度によるものです。プレリリーステストでは、出力速度が秒間 350 トークンを記録しました。
➤ タスクあたりのコストは 2 倍に増加: Gemini 3.5 Flash-Lite のタスクあたりの平均コストは、0.04 ドルから 0.09 ドルへと上昇しました。これは、1M トークンあたり入力 0.30 ドル・出力 2.50 ドルという新価格設定によるものです(Gemini 3.1 Flash-Lite の 0.25 ドル/1.50 ドルから値上げ)。出力トークンの使用量が 1 タスクあたり平均 2 万から 1 万 3 千に減っているにもかかわらず、このコスト増は避けられませんでした。
主要なモデル詳細:
➤ コンテキストウィンドウ: 両モデルとも、前世代と同じ 1M トークンのコンテキストウィンドウを維持しています。
➤ マルチモーダル対応: 両モデルとも、テキスト・画像・動画・音声の入力をサポートし、出力はテキストのみとなります。
➤ 価格設定: Gemini 3.6 Flash の価格は、1M トークンあたり入力 1.50 ドル・出力 7.50 ドルで、Gemini 3.5 Flash(入力 1.50 ドル・出力 9.00 ドル)より安くなりました。一方、Gemini 3.5 Flash-Lite は 1M トークンあたり入力 0.30 ドル・出力 2.50 ドルで、すべての入力モダリティに対して同一の価格設定です。これは Gemini 3.1 Flash-Lite(入力 0.25 ドル・出力 1.50 ドル、音声入力は 0.50 ドル)からの値上げとなります。両モデルとも、キャッシュされた入力トークンには引き続き 90% の割引が適用されます。


原文を表示
Google has released Gemini 3.6 Flash and Gemini 3.5 Flash-Lite. Both halve time per task relative to their predecessors and increase token efficiency, Gemini 3.5 Flash-Lite improves by 11 Intelligence Index points while Gemini 3.6 Flash does not improve in intelligence over 3.5 Flash
Google DeepMind has released the latest updates to the Gemini model family with two new models. We benchmarked Gemini 3.6 Flash and Gemini 3.5 Flash-Lite ahead of release across Intelligence, Time per Task, and Cost per Task

Key takeaways for Gemini 3.6 Flash (high reasoning):
➤ Maintains the same Intelligence as Gemini 3.5 Flash: Gemini 3.6 Flash scores 50 on the Artificial Analysis Intelligence Index, matching Gemini 3.5 Flash, and just below recently released models Muse Spark 1.1 (xhigh, 51) and GPT-5.6 Luna (max, 51). Compared to Gemini 3.5 Flash, Gemini 3.6 Flash maintains similar scores across the Index, with an improvement in GDPval-AA v2 (1421, +72) and a slight regression in HLE (38%, -3 points)
➤ Half the Time per Task: Gemini 3.6 Flash records an average time per task of 1.3 minutes, a more than 50% reduction compared to Gemini 3.5 Flash (2.7). This is driven by increased token efficiency and faster output, with speeds measured at 304 output tokens per second in our pre-launch testing
➤ Slightly lower Cost per Task: Gemini 3.6 Flash’s cost per task decreases ~18%, from $0.59 to $0.50. This is driven by lower output token use and new pricing of $1.50/$7.50 per 1M input/output tokens, down from $1.50/$9.00 for Gemini 3.5 Flash

Key takeaways for Gemini 3.5 Flash-Lite (high reasoning):
➤ Significant Intelligence improvements over Gemini 3.1 Flash Lite: Gemini 3.5 Flash-Lite scores 36 on the Artificial Analysis Intelligence Index, up 11 points from Gemini 3.1 Flash-Lite (25). This places it behind models such as Nemotron 3 Ultra (38) and DeepSeek V4 Flash (max, 40), and above Mistral Medium 3.5 (30). The biggest intelligence gains compared to Gemini 3.1 Flash-Lite are in agentic evaluations, with improvements in GDPval-AA v2 (1140, +498), TerminalBench v2.1 (53.6, +22.5 points) and Tau3-Banking (16.5%, +7.8 points)
➤ Nearly half the Time per Task: Gemini 3.5 Flash-Lite records an average time per task of 0.6 minutes, nearly a 50% reduction compared to Gemini 3.1 Flash-Lite (1.0). This is driven by increased token efficiency and fast output speed, measured at 350 output tokens per second in our pre-launch testing
➤ More expensive with 2x Cost per Task: Gemini 3.5 Flash-Lite’s average cost per task increases from $0.04 to $0.09, driven by new pricing of $0.30/$2.50 per 1M input/output tokens, up from $0.25/$1.50 for Gemini 3.1 Flash-Lite. This cost increase comes despite using fewer output tokens, falling from 20k to 13k average output tokens per task
Key model details:
➤ Context window: Both models retain the same 1M context window as their predecessors
➤ Multimodality: Both models have text, image, video, and speech input with text output only
➤ Pricing: Gemini 3.6 Flash is priced at $1.50/$7.50 per million input/output tokens, down from Gemini 3.5 Flash at $1.50/$9.00. Gemini 3.5 Flash-Lite is priced at $0.30/$2.50 per million input/output tokens, with the same input pricing across all input modalities. This is an increase from Gemini 3.1 Flash-Lite, which is priced at $0.25/$1.50 per million input/output tokens, with input audio tokens at $0.50. Both models retain the same 90% discount for cached input tokens


AI算出
主要ニュースainew評価標準
AI モデルの新規発表であり、ベンチマーク数値やコスト変化など具体的な技術的・経済的インパクトが含まれているため新規性は高い。ただし日本固有の情報や企業事例は含まれていないため、日本の関連性は低めとなる。
6つの評価軸を見る
- AI関連度
- 100
- 情報源の信頼性
- 25
- 新規性
- 75
- 調べる価値
- 100
- 重複の少なさ
- 100
- 日本での有用性
- 25
関連記事
News to Guide
ニュースの次に確認する
発表内容を、現在の料金や仕様と照らし合わせられる関連ガイドです。
今日のまとめ
AIデイリーブリーフで今日の重要ニュースをまとめ読み