アント傘下「Ling 3.0 Flash」が AI Gateway に登場
Ant Group の「Ling 3.0 Flash」モデルが Vercel AI Gateway で利用可能となり、特にエージェンタワークフロー向けに最適化された MoE アーキテクチャと無料期間の提供により、開発現場での実用性が大幅に向上した。
キーポイント
高性能な MoE モデルの登場
124B パラメータ中約 5.1B をアクティブにする Mixture-of-Experts (MoE) アーキテクチャを採用し、256K トークンのコンテキストウィンドウと思考/非思考モードを備えている。
エージェンタワークフローへの最適化
トークン効率、レイテンシ、コストの制約下で多段階のエージェント実行を行うことを目的として設計されており、コーディングやドキュメント処理などの高頻度ワークフローに特化している。
AI Gateway での無料利用期間
8 月 3 日まで 3 週間限定で無料で利用可能であり、SDK では inclusionai/ling-3.0-flash-free というモデル名を指定することでアクセスできる。
Vercel AI Gateway の機能と価格
プロバイダー料金のそのままの反映(マークアップなし)や BYOK 対応、カスタムレポート、ゼロデータ保持などの機能を備え、開発者のコスト管理と運用効率を支援する。
重要な引用
Ling 3.0 Flash is built for token-efficient agentic inference at production scale
The model is free to use for the next three weeks, through August 3rd.
AI Gateway reflects provider pricing with no markup and does not charge a platform fee on inference
影響分析・編集コメントを表示
影響分析
このニュースは、Vercel の AI Gateway というインフラ上で Ant Group の最先端モデルが即座に利用可能になることを示しており、特にエージェンタベースの開発や大規模コンテキスト処理を必要とする現場にとって、コストとパフォーマンスのバランスを改善する重要な選択肢を提供します。無料期間の設定により、開発者はリスクなく新モデルの実力を検証し、本番環境への導入判断を下すことが容易になります。
編集コメント
エージェンタ AI の実用化において、コスト効率とトークン制限の克服は最大の課題の一つですが、Ling 3.0 Flash はその両方を解決する可能性を秘めた注目すべきモデルです。Vercel AI Gateway を通じた無料提供により、開発者がすぐに検証・導入できる環境が整ったことは、業界全体のスピードアップに寄与する動きと言えます。
アリババグループ傘下の Ant Group が開発した「Ling 3.0 Flash」が、AI Gateway で利用可能になりました。
このモデルは、8月3日まで3週間限定で無料で使用できます。
Ling 3.0 Flash は、総パラメータ数124B(1,240億)を備えつつ、トークンあたり約5.1B(51億)のアクティブなパラメータのみを使用する「Mixture-of-Experts(MoE:混合専門家)」モデルです。コンテキストウィンドウは256K トークンをサポートし、思考モードと非思考モードの両方で動作します。
このモデルは、生産規模でのトークン効率に優れたエージェント推論を目的として設計されています。複数のステップからなるエージェント実行において、より少ないトークン数、低遅延、コスト削減を実現しながら、多くの処理をこなすことができます。主なターゲットは、高頻度で発生するエージェントワークフロー、コード生成エージェント、ドキュメント処理、そして長いコンテキストを必要とする多段階の対話です。
AI SDK を使用して Ling 3.0 Flash を利用するには、モデル名に「inclusionai/ling-3.0-flash-free」を設定してください。
AI Gateway は、モデル呼び出しの一元化、利用状況とコストの追跡、リトライやフェイルオーバーの設定、プロバイダー以上の稼働率を実現するためのパフォーマンス最適化などを行うための統一 API を提供しています。さらに、カスタムレポート機能、データ保持ゼロ(Zero Data Retention)への対応、API キーごとの予算管理、ルーティングルールなどの機能を標準で備えています。
AI Gateway では、プロバイダーの価格をそのまま反映し、マージンを上乗せしません。また、推論コストや「Bring Your Own Key (BYOK):自作キーの使用」リクエストに対してもプラットフォーム利用料は徴収されません。
モデルプレイグラウンドで Ling 3.0 Flash をお試しください。
続きを読む
原文を表示
Ling 3.0 Flash from Ant Group is now available on AI Gateway.
The model is free to use for the next three weeks, through August 3rd.
Ling 3.0 Flash is a Mixture-of-Experts model with 124B total parameters and about 5.1B active per token. It has a 256K token context window and runs in thinking and non-thinking modes.
Ling 3.0 Flash is built for token-efficient agentic inference at production scale, doing more work within tighter token, latency, and cost budgets across multi-step agent runs. The model targets high-frequency agentic workflows, coding agents, document work, and long-context multi-turn interactions.
To use Ling 3.0 Flash, set model to inclusionai/ling-3.0-flash-free in the AI SDK:
AI Gateway provides a unified API for calling models, tracking usage and cost, and configuring retries, failover, and performance optimizations for higher-than-provider uptime. It includes built-in custom reporting, Zero Data Retention support, budgets for API keys, routing rules, and more.
AI Gateway reflects provider pricing with no markup and does not charge a platform fee on inference, including on Bring Your Own Key (BYOK) requests.
Try Ling 3.0 Flash in the model playground.
Read more
関連記事
今日のまとめ
AI日報で今日の重要ニュースをまとめ読み