AI Gateway で GLM 5.2 Fast が Wafer を経由して利用可能に
本文の状態
日本語全文を表示中
詳細モードで約1分の本文を読めます。
同じ出来事の情報源
この情報源を基点に整理
Vercel Blog
Vercel は、Zai の提供する GLM 5.2 Fast モデルを AI Gateway 上で Wafer を介して提供開始した。ベンチマークによると、サーバーレス環境でのスループットは他社より 2 倍高く、小・大コンテキストともに高速な生成を実現している。
Continue in AI NEW LAB
このニュースを、実務の判断につなげる
AI NEW LABで、試したことや先に確認したい条件を共有できます。まずはログインなしで読めます。
AI NEW LABで論点を見るSource Article
元記事を日本語で読む
本文に関係しない購読案内、埋め込み通知、サイト内プロモーションは除いています。
GLM 5.2 Fast via Wafer が AI Gateway で利用可能になりました。
小コンテキスト、大コンテキスト、ツール呼び出しのシナリオにわたる独自のベンチマークに基づくと、サーバーレス環境で GLM-5.2 を提供する他のプロバイダーと比較して、Wafer は 2 倍の高いスループットを実現し、小・大コンテキストケースにおける持続的な生成において、デコード速度とエンドツーエンドの速度でリードしています。
テスト結果では、Wafer 上の GLM 5.2 Fast は以下のパフォーマンスを記録しました:
小コンテキスト: 170+ トークン/秒
大コンテキスト: 200+ トークン/秒
GLM 5.2 Fast を利用するには、AI SDK でモデルを zai/glm-5.2-fast に設定してください。
AI Gateway は、モデル呼び出しの統一 API、使用状況とコストの追跡、リトライ・フェイルオーバー・パフォーマンス最適化の設定を提供し、プロバイダー単体よりも高い稼働率を実現します。組み込みのカスタムレポート機能、ゼロデータ保持(Zero Data Retention)サポート、API キーごとの予算管理など、多くの機能を備えています。
AI Gateway はプロバイダーの価格をそのまま反映し、マージンを加算せず、推論時にもプラットフォーム料金を徴収しません。Bring Your Own Key (BYOK) リクエストにも同様に適用されます。
モデルプレイグラウンドで GLM 5.2 Fast をお試しください。
続きを読む
原文を表示
GLM 5.2 Fast via Wafer is now available on AI Gateway.
Based on our own benchmarking across small-context, large-context, and tool-call scenarios, Wafer delivers a 2x higher throughput than other providers serving GLM-5.2 on serverless, leading on decode and end-to-end speed for sustained generation in the small- and large-context cases.
In our testing, GLM 5.2 Fast on Wafer measured:
Small context: 170+ tok/s
Large context: 200+ tok/s
To use GLM 5.2 Fast, set model to zai/glm-5.2-fast in the AI SDK:
AI Gateway provides a unified API for calling models, tracking usage and cost, and configuring retries, failover, and performance optimizations for higher-than-provider uptime. It includes built-in custom reporting, Zero Data Retention support, budgets for API keys, and more.
AI Gateway reflects provider pricing with no markup and does not charge a platform fee on inference, including on Bring Your Own Key (BYOK) requests.
Try GLM 5.2 Fast in the model playground.
Read more
関連記事
News to Guide
ニュースの次に確認する
発表内容を、現在の料金や仕様と照らし合わせられる関連ガイドです。
今日のまとめ
AIデイリーブリーフで今日の重要ニュースをまとめ読み