OpenAI、新機能「Presence」を発表
OpenAI は、企業向けに信頼性の高い AI エージェントを本番環境で運用するための「Presence」製品を発表し、ポリシーやガードレールを組み合わせた実用的な導入フレームワークを提供した。
キーポイント
本番環境での信頼性確保
AI エージェントが単に動作するだけでなく、高価値な業務を安定的に遂行できるよう、システム評価と展開専門知識を強化した製品である。
役割特化型導入とポリシー制御
各デプロイメントで特定のジョブ(例:請求処理、IT サポート)に限定し、エージェントの権限や承認フロー、人間へのエスカレーションルールを企業が設定可能。
継続的な学習と改善プロセス
本番運用中のセッションデータやエスカレーション事例からギャップを特定し、Codex を活用してエージェントの能力を安全に更新・適応させる仕組みを持つ。
重要な引用
The challenge for enterprises is no longer proving that AI agents can work, it’s making them reliable enough to do high-value work in production.
Presence pairs model reasoning with policies, guardrails, and escalation rules that verify accuracy and performance.
Codex proposes updates that teams can test and approve, helping the agent adapt as customer behavior changes.
影響分析・編集コメントを表示
影響分析
この発表は、AI エージェントが実験段階から実社会の複雑な業務に深く統合されるための重要なマイルストーンとなる。OpenAI が提供する包括的な管理・評価ツールにより、企業がリスクを許容範囲内に抑えながら AI の導入を加速できる環境が整った。
編集コメント
「Presence」という名称が示す通り、AI エージェントを単なるチャットボットではなく、企業のインフラに深く組み込まれた信頼できるパートナーとして位置づける戦略が明確です。特に Codex を活用した継続的な改善プロセスは、実運用における最大の課題である「変化への適応」に対する具体的な回答と言えます。
企業が直面する課題は、もはや AI エージェントが動作することを証明することではありません。重要なのは、生産環境において高価値な業務を信頼して実行できるレベルにまで高めることです。製品やポリシー、ユーザーの行動は常に変化するため、エージェントの振る舞いもそれに応じて適応する必要があります。これを実現するには、単なるモデルだけでは不十分です。条件の変化に合わせてエージェントを改善しつつ、コントロールを手放さないためのシステム、評価手法、そして展開に関する専門知識が不可欠です。
私たちは「OpenAI Presence」を発表します。これは、信頼できる AI エージェントの導入を支援する、実戦で検証済みの製品です。Presence は、質問への回答や問題解決、社内システムの活用、承認されたアクションの実行が可能であり、必要に応じて人間へエスカレーションもできます。大規模な企業顧客との長年の実績に基づき、モデルの推論能力と、精度やパフォーマンスを検証するポリシー、ガードレール、エスカレーションルールを組み合わせました。
各導入では、請求書のトラブル対応、保険請求のサポート、社内の IT サービスリクエスト処理など、特定の業務に焦点を当てます。エージェントには、その業務に必要な知識とシステムアクセス権のみが付与されます。企業が設定するポリシーでは、エージェントが何を実行できるか、いつ承認が必要か、そしていつ人間が引き継ぐべきかが明確になります。ローンチ後は、実際の運用セッションやエスカレーション事例から課題が浮き彫りになります。Codex が改善案を提案し、チームはそれをテストして承認することで、顧客の行動変化に合わせてエージェントを適応させていきます。
OpenAI は各顧客と密接に連携し、高価値のワークフローを特定し、必要な知識やシステムをつなぎ合わせ、権限やポリシーを設定し、エージェントをテストして本番環境へ導入します。展開が拡大しても、OpenAI と選定されたシステムインテグレーターが継続してサポートを行います。
Presence は OpenAI の研究チームと緊密に協力して開発されました。すべての展開から得られた一般的な知見は、継続的な研究や製品開発にフィードバックされ、時間とともに全顧客にとっての製品の改善につながっています。
音声およびチャットエージェント向けに本日利用可能
現在、Presence はカスタマーサポート、アウトバウンドセールス、高リスクな内部ワークフローなど、音声とチャットを横断したリアルタイム体験をサポートしています。例えば、請求に関する問い合わせに対応する際、顧客の要求を理解し本人確認を行い、アカウント情報を検索し、会社の方針を適用して承認されたアクションを実行するといった一連の流れを処理できます。
各企業は、ポリシーや評価基準、エスカレーションルールなど展開全体で統一すべき事項と、ワークフローやチャネルごとに柔軟に変更すべき事項を決定します。これにより、チームはすでに機能している仕組みを基盤に構築し、ゼロから作り直すことなく新しいユースケースへと拡張することが可能になります。
Presence は、本番環境でエージェントを運用するために必要なコンポーネントを一つにまとめました。具体的には、ポリシーと標準作業手順(SOP)、ガードレール、承認済みアクション、シミュレーション機能、評価ツール、そして Codex を活用した改善プロセスです。
これらすべてのコンポーネントを組み合わせることで、チームは自社のシステムと連携し、エージェントの振る舞いを定義し、パフォーマンスを評価し、ポリシーを適用し、ローンチ後の変更管理を行うことができます。

大手企業で実証済み
Presence は、顧客との連携を通じて何年にもわたってエージェントを運用・展開してきた経験に基づき構築された製品です。ミッションクリティカルな環境における厳しい要件に応えるため、大胆に設計されています。これまでに実施されたすべての展開から得られた知見が継続的に OpenAI の研究や製品開発にフィードバックされ、顧客にとってのメリットが複利のように積み上がっています。
Presence は、OpenAI の英語圏向け電話サポート窓口(1-888-GPT‑0090)を稼働させています。このシステムは、自由な形式のリクエストに応じ、通話者の身元を確認し、アカウントの文脈を活用して承認されたアクションを実行します。運用開始から数週間で、我々がフロントラインの人間サポート品質を評価する基準に達し、あるいはそれを上回るパフォーマンスを発揮しました。現在では、問い合わせの 75% を人間の介入なしで解決しています。ローンチチームとの連携により、Codex を活用した改善ループが機能し、わずか 10 日間で人間のハンドオフ(引き継ぎ)を 15 ポイント削減することに成功しました。
大手企業もまた、同じく実証済みの基盤の上に構築を進めています:
- BBVA は、メキシコにおける日常の銀行業務向けに AI を活用した音声サポートを検討しています。
- ソフトバンクは、自然な日本語による顧客との対話テストを行っています。
- IAG は、悪天候など需要が急増するイベント時にタイムリーなサポートを提供できる仕組みを模索しています。
1 of 3
ローンチ前・中・後の信頼構築に特化
Presence の本番環境への展開前に、チームは一般的なリクエストやエッジケース、リスクの高いシナリオに対してテストを実施できます。シミュレーションと評価ツールにより、適切な結果が得られたか、ポリシーに従っていたか、ツールの使用が適切だったか、必要に応じてエスカレーションが行われたかなどを確認します。また、対話が企業の境界を超えた場合にはガードレールが介入して制御を行います。

本番環境へのリリース後も、この作業は止まりません。実際の運用セッションやエスカレーション事例、品質に関するシグナルを分析することで、エージェントがどこでうまく機能し、どこに注力すべきかがわかります。ポリシーや製品、ユーザーの行動の変化に応じてです。Codex は Presence プラグインを活用してこれらのシグナルを検証し、改善案を提案します。チームは本番環境にある既存バージョンに対して各提案された変更点をテストした上で、管理されたロールアウトとして承認します。
Presence は、ビジネスや顧客、従業員が変化するにつれて共に進化し、改善していくように設計されています。その一方で、ビジネスのコントロールは常に維持されます。OpenAI の世界トップクラスの研究成果を基盤に構築されており、モデルや機能の向上に伴い、さらに発展を続けていきます。もし利用ケースが現在の製品サポート範囲を超えた場合でも、OpenAI のフォワードデプロイエンジニア(FDE)とパートナーが顧客と連携し、本番環境での実装をサポートします。

利用可能性
OpenAI Presence は、対象となるエンタープライズ顧客向けに、限定一般提供プログラムを通じてデプロイ済み製品として利用可能です。導入プロジェクトは、OpenAI のフォワード・ディプロイメントエンジニアと、選定されたグローバルなシステムインテグレーターが主導します。現時点では、セルフサービス型製品としての提供はまだ行われていません。
OpenAI Presence の紹介にあわせて、引き続き OpenAI API を通じて最先端モデルへのアクセスを提供し、音声利用顧客を支援してまいります。
Presence が貴社の組織に適しているかどうかをご検討される場合は、担当の OpenAI アカウントチームまでお問い合わせください。
原文を表示
The challenge for enterprises is no longer proving that AI agents can work, it’s making them reliable enough to do high-value work in production. Agent behavior must also adapt as products, policies, and user behavior change. That requires more than a model: it requires the systems, evaluations, and deployment expertise to improve agents as those conditions change without giving up control.
We’re introducing OpenAI Presence, a battle-tested product that helps enterprises deploy trusted AI agents that can answer questions, resolve issues, use company systems, take approved actions, and escalate to people when needed. Proven through years of working with customers at enterprise-scale, Presence pairs model reasoning with policies, guardrails, and escalation rules that verify accuracy and performance.
Each deployment starts with a specific job, such as resolving billing issues, supporting insurance claims, or resolving employee IT service requests. The agent receives only the knowledge and system access required for that job. The company sets the policies: what the agent can do, when it needs approval, and when a person should take over. After launch, production sessions and escalations reveal gaps. Codex proposes updates that teams can test and approve, helping the agent adapt as customer behavior changes.
OpenAI works alongside each customer to identify each high-value workflow, connect the necessary knowledge and systems, establish permissions and policies, test the agent, and bring it into production. As the deployment expands, OpenAI and select systems integrators can continue to support it. Presence was developed in tight collaboration with OpenAI’s Research team, with generalized insights from every deployment informing ongoing research and product development and improving the product for all customers over time.
Available today for voice and chat agents
Today, Presence supports real-time experiences across voice and chat, such as customer support, outbound sales, and high-risk internal workflows. A customer might use it to resolve a billing issue—from understanding the request and verifying the customer to looking up account information, applying company policy, and taking an approved action.
Companies decide what remains consistent across deployments—such as policies, evaluations, and escalation rules—and what should change for each workflow or channel. This lets teams build on what works and expand to new use cases without starting over.
Presence brings together the components teams need to run agents in production: policies and standard operating procedures, guardrails, approved actions, simulations, evaluation tools, and a Codex-powered improvement process.
Together, these components help teams connect company systems, define how agents should behave, evaluate performance, enforce policies, and manage changes after launch.

Proven in leading enterprises
Built through years of deploying agents with customers, Presence is an ambitious product shaped by the demands of mission-critical environments. Every deployment generated insights that continuously inform OpenAI research and product development, creating compounding benefits for customers.
Presence powers OpenAI’s English-language phone support channel at 1-888-GPT‑0090, handling open-ended requests, verifying callers, using account context, and taking approved actions. Within weeks, it met or exceeded benchmarks we use to grade frontline human-support quality and now resolves 75% of inbound issues without human assistance. Working with our launch team, its Codex-powered improvement loop reduced human handoffs by 15 percentage points in just 10 days.
Leading enterprises are also building on the same proven foundation:
- BBVA is exploring AI-powered voice support for everyday banking needs in Mexico.
- SoftBank is testing natural Japanese-language customer conversations.
- IAG is exploring timely support during high-demand events such as severe weather.
1 of 3
Built for trust before, during, and after launch
Before a Presence deployment reaches users, teams can test it against common requests, edge cases, and higher-risk scenarios. Simulations and graders check whether it reached the right outcome, followed policy, used tools correctly, and escalated when appropriate. Guardrails can intervene when an interaction moves outside the company’s boundaries.

That work does not stop at launch. Production sessions, escalations, and quality signals can show teams where an agent is working well and where it needs attention as policies, products, and user behavior change. Codex using the Presence plugin investigates those signals and suggests updates. Teams can test each proposed change against the version in production, then approve a controlled rollout.
Presence is also designed to learn with you—improving as your business, customers, and employees evolve while the business stays in control. Built on OpenAI’s world-class research, it continues to advance as our models and capabilities improve. When a use case goes beyond what the product supports today, OpenAI Forward Deployed Engineers (FDEs) and partners can work with the customer to bring it into production.

Availability
OpenAI Presence is available to eligible enterprise customers as a deployed product through a limited general availability program. Deployments are led by OpenAI Forward Deployed Engineers and select global systems integrators. Presence is not yet available as a self-serve product.
As we introduce OpenAI Presence, we'll continue supporting voice customers with access to our frontier models through the OpenAI API.
To explore whether Presence is right for your organization, contact your OpenAI account team.
関連記事
今日のまとめ
AI日報で今日の重要ニュースをまとめ読み