Thinking Machines の軽量モデル「Inkling Small」が AI Gateway に追加
Vercel は Thinking Machines の新モデル「Inkling Small」を AI Gateway に公開し、推論コストと性能のバランスに優れ、音声・画像へのネイティブ対応やゼロデータ保持機能を備えた実用性を高めた。
AI深層分析を開く2026年7月31日 04:32
AI深層分析
キーポイント
高性能と低コストの両立
Inkling Small は大型モデルに匹敵する性能を、サイズが約4分の1で提供し、タスクあたりの計算リソース使用量を大幅に削減している。
マルチモーダル推論能力
音声と画像に対するネイティブな推論機能を備え、プログラムによる画像の切り取りや拡大、詳細検査を可能にしてドキュメント分析を支援する。
柔軟な推論制御機能
ユーザーは最小から最大までの推論努力レベルを選択でき、品質とコスト・遅延時間のトレードオフを自在に調整できる。
セキュリティとコスト効率
ゼロデータ保持(Zero Data Retention)に対応し、プロバイダーがリクエスト後にデータを削除する経路のみを経由するため、開発者や企業のデータ保護要件を満たす。
重要な引用
Inkling Small reaches performance comparable to the larger Inkling model at about a quarter of the size, using much less compute per task.
Controllable thinking effort lets you trade quality against cost and latency, from minimal to maximum reasoning.
For visual tasks, it can crop, zoom, and inspect images programmatically.
編集コメントを表示
編集コメント
Thinking Machines の新モデルは、推論コストを削減しつつ高度な視覚・音声処理能力を維持する点で実用的な価値が高い。Vercel AI Gateway との連携により、セキュリティ要件を満たす形で迅速に導入できるため、開発現場での採用が期待される。
Source Article
元記事を日本語で読む
本文に関係しない購読案内、埋め込み通知、サイト内プロモーションは除いています。
Thinking Machines 社の「Inkling Small」が AI Gateway で利用可能になりました。
Inkling Small は、より大規模な Inkling モデルと同等の性能を、サイズが約4分の1で実現します。また、タスクあたりの計算リソース消費も大幅に削減されています。音声や画像に対するネイティブな推論能力を持つ汎用モデルであり、推論、エージェントによるコーディング、ツール使用などの分野でも高い評価を得ています。
「思考の努力」を制御可能にすることで、コストとレイテンシとのトレードオフを調整できます。最小限から最大限まで、必要に応じて推論レベルを選べます。
視覚タスクにおいては、プログラムによって画像の切り抜き、ズーム、詳細な検査を行うことができます。これにより、重要な情報が微小な領域にあるドキュメントやチャートの処理に特に有効です。
Inkling を利用するには、AI SDK でモデルを thinkingmachines/inkling-small に設定してください。
Inkling-Small は「ゼロデータ保持(Zero Data Retention)」にも対応しています。ダッシュボードからチーム全体でオンにするか、リクエストごとに zeroDataRetention: true を指定することで有効化できます。AI Gateway は、この設定が有効な場合のみ、リクエスト後にプロンプトとレスポンスを削除するプロバイダーにルーティングします。
また、Inkling-Small はコーディングやツール使用のワークフローにおいてコスト効率の高い選択肢です。vercel ai-gateway coding-agents setup を実行してコーディングエージェントを AI Gateway に接続し、エージェントの設定で thinkingmachines/inkling-small を選択してください。詳細はコーディングエージェントガイドをご覧ください。
AI Gateway はプロバイダーの価格をそのまま反映し、マークアップを加えません。推論に対してプラットフォーム手数料も徴収しません(Bring Your Own Key (BYOK) リクエストを含む)。モデルプレイグラウンドで Inkling Small をお試しください。
続きを読む
原文を表示
Inkling Small from Thinking Machines is now available on AI Gateway.
Inkling Small reaches performance comparable to the larger Inkling model at about a quarter of the size, using much less compute per task. It is a broad generalist with native reasoning over audio and images, and it holds up well on reasoning, agentic coding, and tool use. Controllable thinking effort lets you trade quality against cost and latency, from minimal to maximum reasoning.
For visual tasks, it can crop, zoom, and inspect images programmatically, which helps on documents and charts where the relevant detail is small.
To use Inkling, set model to thinkingmachines/inkling-small in the AI SDK:
Inkling-Small is compatible with Zero Data Retention. Turn it on team-wide from the dashboard, or per request with zeroDataRetention: true, and AI Gateway routes only to providers that delete prompts and responses after each request.
Inkling-Small is also a cost-efficient choice for coding and tool-use workflows. Run vercel ai-gateway coding-agents setup to connect your coding agents to AI Gateway, then select thinkingmachines/inkling-small in the agent's model configuration. See the coding agents guide.
AI Gateway reflects provider pricing with no markup and does not charge a platform fee on inference, including on Bring Your Own Key (BYOK) requests. Try Inkling Small in the model playground.
Read more
AI算出
主要ニュースainew評価高い
AI モデルの機能強化とプラットフォームへの統合を報じており、新規性が高く検索意図も明確である。ただし、日本固有の導入事例や規制情報は含まれていないため、日本の関連性は低めとなる。
6つの評価軸を見る
- AI関連度
- 100
- 情報源の信頼性
- 100
- 新規性
- 75
- 調べる価値
- 75
- 重複の少なさ
- 100
- 日本での有用性
- 25
関連記事
今日のまとめ
AIデイリーブリーフで今日の重要ニュースをまとめ読み