xAI、画像生成モデル「Imagine Image 2.0」を公開しArenaベンチでOpenAIに次ぐ結果
本文の状態
日本語全文を表示中
詳細モードで約3分の本文を読めます。
同じ出来事の情報源
この情報源を基点に整理
The Decoder
xAI は画像生成モデル「Imagine Image 2.0」を公開し、編集機能やテンプレートを強化した結果、Arena ベンチマークで OpenAI の GPT-Image-2 に次ぐ世界第2位を獲得した。
Continue in AI NEW LAB
このニュースを、実務の判断につなげる
AI NEW LABで、試したことや先に確認したい条件を共有できます。まずはログインなしで読めます。
AI NEW LABで論点を見るAI深層分析を開く2026年8月8日 17:46
AI深層分析
キーポイント
ベンチマークでの高評価獲得
xAI によると、Imagine Image 2.0 は画像編集およびテキストから画像生成の両分野で OpenAI の GPT-Image-2 に次ぐスコアを記録し、競合他社モデルを上回っている。
高度な編集機能の追加
「Magic Wand」による選択領域修正やセグメンテーション、背景除去機能に加え、複数画像を組み合わせたマルチリファレンス編集やスマートリサイズ機能が実装された。
ワークフロー効率化ツールの提供
写真編集からゲームアセット作成までをカバーするプリ設定テンプレートが導入され、反復的な作業フローをサポートする環境が整えられた。
マルチ画像編集とスマートリサイズ機能
最大5枚の画像を組み合わせて1枚の生成を行う「Multi-Ref Editing」や、既存の画像を任意のアスペクト比に変換する「Smart Resize」が追加された。
テンプレートと動画制作の前段階機能
写真編集からゲームアセットまで多様なワークフローをまとめたテンプレートが提供される。また、キャラクターや背景などの要素を別々に生成しながら視覚的スタイルを統一する機能により、フル動画制作への道筋が示された。
重要な引用
xAI has launched Imagine Image 2.0 as a new "Quality Mode" on grok.com/imagine and in Grok's iOS and Android apps.
Imagine 2.0 is designed to follow instructions with fine-grained accuracy, keep typography and layout clean in complex visuals, and stay consistent across multiple generations
In the Arena leaderboards as of August 7, 2026, the faster "low" variant of the model takes second place globally
xAI is also showing a feature that generates characters, locations, and props separately while keeping the visual style consistent across all images.
編集コメントを表示
編集コメント
xAI は画像生成モデルの性能向上に成功し、OpenAI との競争において明確な存在感を示した。特に編集機能とテンプレートによるワークフロー支援は、実務利用における採用意欲を高める要素となっている。
Source Article
元記事を日本語で読む
本文に関係しない購読案内、埋め込み通知、サイト内プロモーションは除いています。
xAI、編集ツールと事前設定済みテンプレートを備えた「Imagine Image 2.0」を公開。Arena ベンチマークでは OpenAI の GPT-Image-2 に次ぐ結果に。
xAI は、grok.com/imagine および Grok の iOS と Android アプリで、新しい「Quality Mode」として Imagine Image 2.0 をリリースしました。同社によると、API アクセスもまもなく提供される予定です。
同社は、Imagine 2.0 が指示を細部まで正確に実行し、複雑なビジュアルでもタイポグラフィやレイアウトを清潔に保ち、複数回の生成間での一貫性を維持できると説明しています。
2026 年 8 月 7 日時点の Arena リーダーボード では、より高速な「low」版が両カテゴリで世界第 2 位を獲得しました。Image Edit Arena では Elo レーティング 1,439 を記録し、OpenAI の GPT-Image-2(1,463)に次ぐ成績です。Text-to-Image Arena でも 1,320 を達成し、同様に GPT-Image-2(1,380)に後れをとっています。

Reve 2.1、Meta の Muse-Image、Alibaba の Qwen-Image-3.0-Pro、Google の Gemini、ByteDance の SeedDream はさらに下位にランクインしています。また、この新モデルはこれらのテストにおいて、前作の「Quality」版を大きく上回る結果を残しました。
反復ワークフロー向けに設計された編集ツール
Imagine Image 2.0 には、いくつかの画像編集機能が搭載されています。xAI によると、「マジックワンド」と呼ばれるツールは、選択した領域のみを変更可能にします。セグメンテーション機能を使えば、ユーザーは正確な範囲を指定できますし、背景除去ツールでは被写体を透過背景付きで書き出せます。
Multi-Ref Editing(マルチリファレンス編集)では、最大 5 つの画像を組み合わせて 1 つの生成結果を作れます。「スマートリサイズ」機能を使えば、既存の画像を任意のアスペクト比に変換でき、不足する部分はモデルが自動的に補完します。
テンプレートと動画制作前の企画立案
xAI はまた、一般的な画像ワークフローを事前設定されたテンプレートにまとめる機能を追加しました。カテゴリは写真編集、商品撮影、マーケティング素材、デザインツール、ゲームアセット、ストリーミング用絵文字など多岐にわたります。
さらに同社は、キャラクター・場所・小道具を別々に生成しながらも、すべての画像で視覚的なスタイルを一貫させる機能も公開しました。これは完全な動画制作ワークフローへの架け橋として位置づけられています。

過剰な hype を排した AI ニュース – 人間が厳選
THE DECODER に登録すれば、広告なしでの閲覧、週刊 AI ニュレター、年 6 回の独占「AI Radar」フロンティアレポート、アーカイブへの完全アクセス、そしてコメント欄の利用が可能になります。
購読はこちらから
https://the-decoder.com/subscription/
原文を表示
xAI releases Imagine Image 2.0 with editing tools and preconfigured templates. The model lands just behind OpenAI's GPT-Image-2 in Arena benchmarks.
xAI has launched Imagine Image 2.0 as a new "Quality Mode" on grok.com/imagine and in Grok's iOS and Android apps. API access for Imagine Image 2.0 is coming soon, according to xAI.
Imagine 2.0 is designed to follow instructions with fine-grained accuracy, keep typography and layout clean in complex visuals, and stay consistent across multiple generations, the company says.
In the Arena leaderboards as of August 7, 2026, the faster "low" variant of the model takes second place globally in both categories. It scores an Elo rating of 1,439 in the Image Edit Arena, behind OpenAI's GPT-Image-2 at 1,463. In the Text-to-Image Arena, it hits 1,320, again trailing GPT-Image-2 at 1,380.

Reve 2.1, Meta's Muse-Image, Alibaba's Qwen-Image-3.0-Pro, Google's Gemini, and ByteDance's SeedDream all rank further down the list. The new model also beats the "Quality" variant of its predecessor by a wide margin in these tests.
Editing tools built for iterative workflows
Imagine Image 2.0 ships with several editing features. A tool called "Magic Wand" modifies only the selected area of an image, according to xAI. A segmentation feature lets users pick precise regions, and a background removal tool exports subjects with a transparent background.
Multi-Ref Editing lets users combine up to five input images into a single generation. "Smart Resize" converts an existing image to any aspect ratio, with the model filling in the extra space on its own.
Templates and video pre-production planning
xAI is also adding templates that bundle common image workflows into preconfigured starting points. Categories span photo editing, product photography, marketing materials, design tools, game assets, and streaming emojis.
xAI is also showing a feature that generates characters, locations, and props separately while keeping the visual style consistent across all images. The company positions this as a stepping stone toward full video production workflows.

AI News Without the Hype – Curated by Humans
Subscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive "AI Radar" frontier report six times a year, full archive access, and access to our comment section.
関連記事
今日のまとめ
AIデイリーブリーフで今日の重要ニュースをまとめ読み