Google の Gemini Omni ビデオモデルが I/O デビュー前に登場、チャット内で動画編集機能を統合
本文の状態
日本語全文を表示中
詳細モードで約2分の本文を読めます。
同じ出来事の情報源
この情報源を基点に整理
TLDR AI
Google は次期イベント「I/O」に先駆け、チャット内で動画のリミックスや編集を直接行える新モデル「Gemini Omni video model」を発表した。このモデルは透かし除去や物体の差し替えなどの編集能力に優れるが、ByteDance の Seedance 2 に比べると映画のような画質では劣る。今後は Flash や Pro といった階層版として展開され、多様なモダリティを Gemini で統一する戦略の一環となる見込みである。
Source Article
元記事を日本語で読む
本文に関係しない購読案内、埋め込み通知、サイト内プロモーションは除いています。
Google の次期 Gemini Omni ビデオモデル に関する新たな兆候が週末に表面化し、Reddit ユーザーが改訂された Gemini インターフェースのスクリーンショットを投稿しました。そこには新しいモデルカードが表示され、「Gemini Omni で創作:新ビデオモデルをご紹介します。動画をリミックスしたり、チャット内で直接編集したり、テンプレートを使ったりできます」という説明文が含まれており、来週のデベロッパーイベントに向けて Google が準備してきたと長年噂されてきた統合アプローチが確認されたようです。この展開は偶発的なものか、あるいは限定された A/B テストの一部の可能性があります。
モデルカードとともに、ユーザーたちは設定内に新しい利用制限タブを発見し、複数の報告によるとビデオ生成でクレジットが急速に消費されるため、Google が Gemini の各表面でテストしているメーター型システム(metered system)と類似した仕組みを示唆しています。初期の出力に対する反応は賛否両論でした。純粋な生成忠実度においては、Omni は ByteDance の Seedance 2 に劣っているように見え、視聴者からは映画のような品質が現在のベンチマークリーダーに一歩譲ると指摘されています。一方、このモデルが目立っていたのは編集機能です。ウォーターマークの除去、クリップ内のオブジェクトの置換、チャットによる指示でのシーン書き換えなど、初めての公開試作としては異例なほどにうまく機能しました。
このパターンは、Gemini でネイティブ画像モデルとして登場し、初期の生成スコアは平凡だったものの編集分野のリーダーボードを制覇し、後に最先端の画像システムへと進化させた Nano Banana と同様のものです。 Google は動画においても同じ戦略を採用しており、発表時の純粋な品質でのトップ地位よりも、Gemini におけるモダリティの統一を優先しているようです。また、Omni は Flash と Pro のような階層化されたバリアントとして提供される可能性があり、現在流通している出力は最もおそらく Flash バージョンからのものであると推測されます。
このタイミングは、5 月 19 日と 20 日に開催される Google I/O に完璧に合致しており、同社は過去にもここで最も野心的な AI の転換点を発表する記録を持っています。イベント直前の短い期間と統制されたリークにより、Google は基調講演の前に反応を集め、物語を形成するための余地を得ています。
原文を表示
Fresh signals around Google’s upcoming Gemini Omni video model surfaced over the weekend, with Reddit users posting screenshots of a revised Gemini interface exposing the new model card. The description read “Create with Gemini Omni: meet our new video model, remix your videos, edit directly in chat, try templates, and more,” appearing to confirm the long-rumored unified approach Google has been preparing ahead of next week’s developer event. The rollout looked either accidental or part of a limited A/B test.
Alongside the model card, users spotted a new usage limits tab inside settings, and several reported that video generation burned through credits fast, hinting at a metered system similar to what Google has been testing across Gemini surfaces. Early outputs drew mixed reactions. On raw generation fidelity, Omni appears to lag behind ByteDance’s Seedance 2, with viewers noting that the cinematic quality is a step behind the current benchmark leader. Where the model stood out was in editing: removing watermarks, swapping objects within clips, and rewriting scenes via chat instructions all worked unusually well for a first public glimpse.
That pattern mirrors Nano Banana, which launched as a native image model on Gemini, debuted with middling generation scores but topped editing leaderboards, and was later upgraded into a frontier image system. Google appears to be running the same playbook for video, prioritizing modality unification under Gemini over raw quality leadership at launch. There are also hints that Omni will ship in tiered variants, likely Flash and Pro, with the outputs circulating now most likely coming from the Flash tier.
The timing fits neatly with Google I/O on May 19 and 20, where the company has a track record of unveiling its most ambitious AI shifts. A short pre-event window paired with a controlled leak gives Google room to gather reactions and shape the narrative before the keynote.
関連記事
今日のまとめ
AIデイリーブリーフで今日の重要ニュースをまとめ読み