Mistral、EU データ処理と優先アクセスを提供開始も制限あり
本文の状態
日本語全文を表示中
詳細モードで約6分の本文を読めます。
同じ出来事の情報源
この情報源を基点に整理
The Decoder
Mistral は欧州および米国でのデータ処理を可能にするリージョン別推論と、混雑時の優先処理を提供するが、機能制限や管理データの地域外処理など重要な制約が存在する。
Continue in AI NEW LAB
このニュースを、実務の判断につなげる
AI NEW LABで、試したことや先に確認したい条件を共有できます。まずはログインなしで読めます。
AI NEW LABで論点を見るAI深層分析を開く2026年8月12日 19:41
AI深層分析
キーポイント
リージョン別推論の開始とコスト
Mistral は欧州(api.eu.mistral.ai)および米国(api.us.mistral.ai)のエンドポイントを一般公開し、データ処理を地域内に留める機能を導入した。ただし、標準価格に 10% の追加料金が発生する。
機能とモデルの制限
リージョン別エンドポイントでは関数呼び出しのみが利用可能で、エージェント、バッチ処理、ファイル管理といったステートフルな機能は提供されない。また、利用可能なモデル一覧は地域ごとに異なり、固定リストは公開されていない。
データ主権の範囲と限界
計算ステップのみが指定された地域で行われるため、アカウント設定や API キー、請求情報などは依然として地域外で処理される可能性がある。完全なデータ主権を保証するには「ゼロデータ保持」設定の確認が必要である。
プライオリティティアの提供
混雑時にレスポンス時間を短縮する「Priority Tier(優先ティア)」がベータ版として展開され、99.5% の稼働率 SLA が契約で保証されている。
優先ティアのサービスレベルと料金
このティアは月間約3.5時間のダウンタイムを許容する99.5%の稼働率SLAを提供し、標準価格の1.75倍(75%の追加料金)で課金される。
重要な引用
Regional routing costs 10 percent on top of standard pricing.
Account settings, API keys, billing, and usage stats can still be processed outside the chosen region.
What's regional is the compute step, not the whole platform.
Setting it to "auto" sends the request through the fast lane when capacity is available.
編集コメントを表示
編集コメント
データ主権を重視する欧州の企業にとって、機能制限が明確に示された今回の発表は実装計画の見直しを迫る重要な情報となる。完全な地域内処理を求める場合は、ステートフル機能の利用可否や追加コストを慎重に評価する必要がある。
Source Article
元記事を日本語で読む
本文に関係しない購読案内、埋め込み通知、サイト内プロモーションは除いています。
「地域別推論」の一般提供が開始されました。顧客は欧州エンドポイント(api.eu.mistral.ai)または米国エンドポイント(api.us.mistral.ai)にリクエストを送信でき、処理はその地域の枠内で完結します。これは、顧客データが EU 域外へ流出しないことを証明する必要がある銀行や政府機関、保険会社にとって重要な機能です。また、ネットワーク経路が短縮されるため、レイテンシの低減も期待できます。
デフォルトのエンドポイントを利用している場合、リクエストがどこで処理されるかについての保証はありません。地域別ルーティングを利用するには、標準価格に 10% の追加料金が発生します。
EU データ処理には大きな制限がある
ただし、プラットフォームのアドオンツールのうち、機能呼び出し(function calling)のみが地域別エンドポイントと連携可能です。これはモデルが外部 API をトリガーする機能を指します。エージェント、バッチ処理、ファイル管理は、地域別アドレスでは利用できません。
モデルの選定も地域によって異なります。Mistral は固定リストを公開しておらず、顧客自身が各エンドポイントをクエリして、そこで利用可能なモデルを確認する必要があります。
その背景には、単純なモデル呼び出しと中間状態の保存との間に大きな隔たりがあることが考えられます。標準的なモデル呼び出しでは永続ストレージは不要ですが、エージェント、バッチジョブ、ファイルストレージは、単一の呼び出しを超えてデータを保持します。具体的には、中間ステップやアップロードされたドキュメントなどが該当します。
Mistral はこれらの機能を「ステートフル(状態を保持する)」と呼んでいます。これらには追加のオンプレミスインフラが必要となる可能性がありますが、同社はその点についてはまだ公式に確認していません。
「主権」という言葉が示す範囲よりも、実際の適用範囲は狭いです。ドキュメントによると、アカウント設定、API キー、請求情報、利用統計などは、選択したリージョン外でも処理されることがあります。ブログ記事では、その地域外の下請け業者への限定的かつ安全なデータ転送にも言及されています。実際には「計算処理」のみが地域限定であり、プラットフォーム全体がそうであるわけではありません。リクエストの保存やログ記録の有無は、「ゼロ・データ・レテンション」と呼ばれる別の設定によって決まります。
つまり実務的には、契約文書を直接モデルに送信する顧客は EU エンドポイントを利用できますが、ファイル管理機能を持つエージェントや Files API を利用したい顧客には同じ保証は適用されません。
企業が高速キューのために支払う理由
2 つ目の提供サービスが「Priority Tier」です。現在はオープンベータ版として提供されています。すべての顧客は同じデータセンターを共有しています。リクエストが集中すると応答時間が遅延しますが、Priority Tier はその混雑時に優先処理される専用レーンのようなものです。有料プランの顧客のリクエストは、通常トラフィックよりも先に処理されます。Mistral が狙っているのは、カスタマーサービスチャットボットや工場の生産システムなど、レイテンシが直接コストに直結するユースケースです。
このティアには 99.5% の稼働率 SLA(サービスレベル契約)が含まれており、これは月間約 3 時間半のダウンタイムを許容する法的に保証されたサービス水準です。一方、Mistral の標準ティアにはこのような保証はありません。
顧客は、service_tier という単一の API パラメータを設定することで優先アクセスを有効化できます。このパラメータに「auto」を指定すると、利用可能な容量がある場合にリクエストが高速パスへ転送されます。デフォルト値は「standard_only」で、こちらは通常の経路を経由します。
各顧客には、1 分あたりの優先処理対象となるリクエスト数について個別に交渉されたレート制限が設定されています。この上限を超えてもリクエスト自体は失敗せず、自動的に標準処理へフォールバックされます。API のレスポンスには実際にどのティアで処理が行われたかが含まれるため、顧客は支払った対価に見合ったサービスを受けられているかを確認できます。
ミストラルの優先アクセス料金は、標準価格の 1.75 倍(75% の上乗せ)です。ただし、プロンプトキャッシングによる割引(繰り返し出現するテキストセグメントを保存し、より低いレートで課金する仕組み)は適用されます。ドキュメントによると、最大 90% に達するこれらの割引がまず計算され、その後に優先アクセスの上乗せ料金が加算される順序です。
なお、この優先ティアはセルフサービスではなく、顧客はミストラルの営業チームと契約を結ぶ必要があります。
ミストラルはまた、他のプロバイダーからのオープンモデルも自社のプラットフォームで提供し始めました。その第1弾として登場したのは、中国のAI企業Z.aiが提供する「GLM-5.2」です。これはミストラル自社モデルと同じ地域ルールと保証の下で動作します。
この計算リソースを賄うために、ミストラルは大手顧客から数年間の購入コミットメントを集約し、「European Compute Units」というパッケージとして提供しています。欧州に新しいデータセンターを建設する事業性が成立するのは、十分な数の企業が長期契約を結んでくれる場合に限られるという考えです。
ミストラルは「Open Secure AI Alliance」およびNvidiaの「Nemotron」コンソーシアムのメンバーでもあります。同社は、サードパーティのモデル重みを自社のプラットフォーム上でホストすることを、これまでの活動の自然な延長と捉えています。
原文を表示
Regional inference is now generally available. Customers can send requests to a European endpoint (api.eu.mistral.ai) or a US endpoint (api.us.mistral.ai), and processing stays in that region. That matters for banks, government agencies, and insurers that need to prove customer data never leaves the EU. Shorter network paths also mean lower latency. Anyone using the default endpoint gets no guarantee about where their request is processed. Regional routing costs 10 percent on top of standard pricing.
EU data processing comes with significant limits
However, among the platform's add-on tools, only function calling works with regional endpoints, meaning the model's ability to trigger external APIs. Agents, batch processing, and file management aren't available at the regional addresses. Model selection varies by region, too. Mistral doesn't publish a fixed list, customers have to query each endpoint to see what's there.
The likely reason is the gap between a simple model query and storing intermediate state. A standard model call needs no persistent storage. Agents, batch jobs, and file storage hold data beyond a single call, things like intermediate steps or uploaded documents. Mistral calls these features "stateful." They probably require extra on-site infrastructure, though the company hasn't confirmed that.
The scope is also narrower than the word "sovereignty" suggests. Account settings, API keys, billing, and usage stats can still be processed outside the chosen region, according to the documentation. The blog post also mentions limited, secured transfers to subcontractors outside the region. What's regional is the compute step, not the whole platform. Whether requests get stored or logged afterward depends on a separate setting called Zero Data Retention.
What this means in practice is that customers who send contract text directly to a model can run it through the EU endpoint. Anyone who needs agents or file management through the Files API won't get the same guarantee.
Why companies would pay for a faster queue
The second offering is the Priority Tier, currently in open beta. All customers share the same data centers. When lots of requests hit at once, response times go up. The Priority Tier is a fast lane: paying customers' requests get processed ahead of regular traffic when things get busy. Mistral is targeting use cases where latency costs real money, like a customer service chatbot or a production system on a factory floor.
The tier includes an uptime SLA of 99.5 percent, a contractually guaranteed service level that allows roughly three and a half hours of downtime per month. Mistral's standard tier has no such guarantee.
Customers activate priority access through a single API parameter called service_tier. Setting it to "auto" sends the request through the fast lane when capacity is available. The default value is "standard_only," which takes the regular path. Each customer also gets individually negotiated rate limits for how many requests per minute get priority treatment. Going over that limit doesn't fail the request. It just falls back to standard processing. The API response shows which tier actually handled the request, so customers can check whether they're getting what they pay for.
Mistral charges 1.75x the standard price, a 75 percent surcharge. Discounts from prompt caching, where repeated text segments are stored and billed at lower rates, still apply. Those discounts can reach 90 percent and get calculated first according to the documentation, with the priority surcharge applied after. The Priority Tier isn't self-service. Customers have to sign a contract with Mistral's sales team.
Third-party models join the platform
Mistral is also opening its platform to open models from other providers. First up is GLM-5.2 from Chinese AI company Z.ai, which runs under the same regional rules and guarantees as Mistral's own models. To fund the compute capacity this requires, Mistral is collecting multi-year purchase commitments from large customers, packaged as European Compute Units. The idea is that building new data centers in Europe only pencils out if enough companies commit long-term.
Mistral is a member of the Open Secure AI Alliance and Nvidia's Nemotron coalition. The company sees hosting third-party model weights on its platform as a natural extension of that work.
関連記事
News to Guide
ニュースの次に確認する
発表内容を、現在の料金や仕様と照らし合わせられる関連ガイドです。
今日のまとめ
AIデイリーブリーフで今日の重要ニュースをまとめ読み