OpenAI、対話型 AI「Presence」を発表
OpenAI は、企業向けに信頼性の高い AI エージェントを本番環境で運用するための製品「Presence」を発表し、ポリシーやガイドラインの統合、自動化された改善プロセスを通じて実用化を加速させる。
キーポイント
本番環境向けの信頼性重視設計
AI エージェントが単に動作するだけでなく、高価値な業務で安定して動作し、ポリシーやユーザー行動の変化に適応できるシステムとして「Presence」を構築した。
ジョブベースの制限とガバナンス
各導入は特定のタスク(請求処理や社内サポートなど)に限定され、必要な知識とアクセス権のみが付与されるため、リスク管理が強化されている。
継続的な学習と改善プロセス
本番運用中の失敗事例やエスカレーションを分析し、Codex が提案する更新案をチームがテスト・承認することで、エージェントの能力を継続的に向上させる。
重要な引用
The challenge for enterprises is no longer proving that AI agents can work, it's making them reliable enough to do high-value work in production.
Presence pairs model reasoning with policies, guardrails, and escalation rules that verify accuracy and performance.
Codex proposes updates that teams can test and approve, helping the agent adapt as customer behavior changes.
影響分析・編集コメントを表示
影響分析
この発表は、AI エージェントが実験段階から実社会での信頼性の高い業務遂行へと移行する重要な転換点を示しています。企業側が完全に制御権を持ちながら、システム的なガバナンスと自動学習機能を活用することで、大規模な導入リスクを低減し、実用化のスピードを劇的に加速させる可能性があります。
編集コメント
「Presence」の登場は、AI エージェントが単なるチャットボットを超え、企業の重要な業務プロセスを担う信頼できるパートナーへと進化しようとしていることを示しています。特に、人間の介入が必要なエスカレーションルールと自動改善ループを組み合わせたアプローチは、実務での導入障壁を下げる画期的な試みと言えるでしょう。
企業が直面する課題は、AI エージェントが機能することを証明することではなく、高価値な業務を本番環境で信頼して実行できるレベルにまで高めることです。製品やポリシー、ユーザーの行動が変化する中で、エージェントの振る舞いも適応する必要があります。そのためには、モデルだけでなく、条件の変化に合わせてエージェントを改善しつつも制御権を放棄しないためのシステム、評価手法、そして展開に関する専門知識が必要です。
私たちは「OpenAI Presence」を発表します。これは、質問への回答、問題解決、社内システムの活用、承認されたアクションの実行、必要に応じて人間へのエスカレーションなどを実現できる信頼できる AI エージェントを企業が導入するための、実戦で検証済みの製品です。大規模企業での顧客との長年の実績に基づき、モデルの推論能力と、精度やパフォーマンスを検証するポリシー、ガードレール、エスカレーションルールを組み合わせています。
各導入は、請求書のトラブル対応、保険請求のサポート、社内の IT サービスリクエスト処理など、特定の業務から始まります。エージェントにはその業務に必要な知識とシステムアクセス権のみが付与されます。企業が設定するのは、エージェントが何を行えるか、いつ承認が必要か、そしていつ人間が引き継ぐべきかなどのポリシーです。ローンチ後、本番環境でのセッションやエスカレーション事例から課題が明らかになります。Codex が改善案を提案し、チームはそれをテストして承認することで、顧客の行動変化に合わせてエージェントを適応させていきます。
OpenAI は各顧客と連携し、価値の高いワークフローを特定し、必要な知識やシステムを接続し、権限やポリシーを設定し、エージェントをテストして本番環境へ導入します。展開が拡大しても、OpenAI と選定されたシステムインテグレーターが継続的にサポートを提供できます。
Presence は OpenAI の研究チームと緊密に協力して開発されました。すべての展開から得られた一般的な知見は、継続的な研究や製品開発に反映され、時間の経過とともに全顧客向けに製品が改善されていきます。
音声およびチャットエージェント向けに本日利用可能
現在、Presence はカスタマーサポート、アウトバウンドセールス、高リスクの内部ワークフローなど、音声とチャットを跨ぐリアルタイム体験をサポートしています。例えば、請求に関する問い合わせに対応する際、顧客からのリクエストを理解し本人確認を行い、アカウント情報を検索し、会社のポリシーを適用して承認されたアクションを実行するといった一連の流れを処理できます。
各企業は、展開全体で統一すべき事項(ポリシー、評価基準、エスカレーションルールなど)と、ワークフローやチャネルごとに柔軟に変更すべき事項を決定します。これにより、チームはすでに機能している仕組みを基盤に構築し、ゼロから作り直すことなく新しいユースケースへと展開することが可能になります。
Presence は、本番環境でエージェントを運用するために必要なコンポーネントを集約しています。具体的には、ポリシーと標準作業手順(SOP)、ガードレール、承認済みアクション、シミュレーション、評価ツール、そして Codex を活用した改善プロセスです。
これらすべてのコンポーネントを組み合わせることで、チームは自社のシステムと連携し、エージェントの行動指針を定義し、パフォーマンスを評価し、ポリシーを適用し、リリース後の変更管理を行うことが可能になります。

大手企業で実証済み
Presence は、顧客との連携を通じて何年にもわたりエージェントを運用・展開してきた経験に基づき構築された、極めて野心的な製品です。ミッションクリティカルな環境における要求に応えるために設計されており、これまでに実施されたすべての展開から得られた知見が継続的に OpenAI の研究および製品開発にフィードバックされています。その結果、顧客にとってのメリットは複利のように積み上がっていきます。
Presence は、OpenAI の英語圏向け電話サポート窓口(1-888-GPT‑0090)を稼働させています。ここでは、自由な形式のリクエストに応じ、通話者の身元確認を行い、アカウントの文脈を活用し、承認されたアクションを実行しています。リリースから数週間以内に、当社のフロントラインの人間によるサポート品質を評価するベンチマークに到達、あるいはそれを上回る性能を発揮。現在では、入ってくる問い合わせの 75% を人間の介入なしで解決できるようになっています。ローンチチームとの連携により、Codex を活用した改善ループが機能し、わずか 10 日間で人間による手動引き継ぎを 15 ポイント削減することに成功しました。
大手企業もまた、同じく実証済みの基盤の上に構築を進めています:
- BBVA は、メキシコにおける日常的な銀行取引のための AI 音声サポートを検討中。
- ソフトバンクは、自然な日本語による顧客との対話テストを実施中。
- IAG(国際航空グループ)は、悪天候など需要が集中するイベントにおいてタイムリーなサポートを提供できるか検討中。
1 of 3
ランチ前・中・後の信頼構築を前提とした設計
Presence の導入前に、チームは一般的なリクエストやエッジケース、リスクの高いシナリオに対してテストを行うことができます。シミュレーションと評価ツールにより、適切な結果が得られたか、ポリシーに従っていたか、ツールの使用が適切だったか、必要に応じてエスカレーションされたかなどを検証します。また、やり取りが企業の境界を越えた際には、ガードレールが介入して制御を行います。

この作業はローンチで終わるものではありません。本番環境でのセッション、エスカレーション事例、品質の指標などから、エージェントがどこでうまく機能し、どこに注意が必要かを把握できます。ポリシーや製品、ユーザー行動の変化に応じてです。Codex は Presence プラグインを利用してこれらの信号を調査し、改善案を提示します。チームは本番環境のバージョンと比較して各提案を検証した上で、管理されたロールアウトを承認します。
Presence は、ビジネス、顧客、従業員の進化に合わせて学習・改善するように設計されていますが、最終的なコントロールは企業が維持します。OpenAI の世界トップクラスの研究成果に基づいて構築されており、モデルや機能の向上に伴いさらに発展していきます。製品で現在サポートしている範囲を超えるユースケースが生じた場合は、OpenAI のフォワードデプロイエンジニア(FDE)とパートナーが顧客と連携し、本番環境での運用を実現します。

利用可能性
OpenAI Presence は、対象となるエンタープライズ顧客向けに限定一般提供プログラムを通じて展開製品として利用可能です。導入プロジェクトは、OpenAI のフォワードデプロイエンジニアと選抜されたグローバルシステムインテグレーターが主導します。現時点ではセルフサービス型製品としての提供はまだ行われていません。
OpenAI Presence の発表に伴い、引き続き OpenAI API を通じて最先端モデルへのアクセスを提供し、音声利用者のサポートを継続していきます。
Presence が貴社の組織に適しているか検討される場合は、OpenAI のアカウントチームまでお問い合わせください。
原文を表示
The challenge for enterprises is no longer proving that AI agents can work, it’s making them reliable enough to do high-value work in production. Agent behavior must also adapt as products, policies, and user behavior change. That requires more than a model: it requires the systems, evaluations, and deployment expertise to improve agents as those conditions change without giving up control.
We’re introducing OpenAI Presence, a battle-tested product that helps enterprises deploy trusted AI agents that can answer questions, resolve issues, use company systems, take approved actions, and escalate to people when needed. Proven through years of working with customers at enterprise-scale, Presence pairs model reasoning with policies, guardrails, and escalation rules that verify accuracy and performance.
Each deployment starts with a specific job, such as resolving billing issues, supporting insurance claims, or resolving employee IT service requests. The agent receives only the knowledge and system access required for that job. The company sets the policies: what the agent can do, when it needs approval, and when a person should take over. After launch, production sessions and escalations reveal gaps. Codex proposes updates that teams can test and approve, helping the agent adapt as customer behavior changes.
OpenAI works alongside each customer to identify each high-value workflow, connect the necessary knowledge and systems, establish permissions and policies, test the agent, and bring it into production. As the deployment expands, OpenAI and select systems integrators can continue to support it. Presence was developed in tight collaboration with OpenAI’s Research team, with generalized insights from every deployment informing ongoing research and product development and improving the product for all customers over time.
Available today for voice and chat agents
Today, Presence supports real-time experiences across voice and chat, such as customer support, outbound sales, and high-risk internal workflows. A customer might use it to resolve a billing issue—from understanding the request and verifying the customer to looking up account information, applying company policy, and taking an approved action.
Companies decide what remains consistent across deployments—such as policies, evaluations, and escalation rules—and what should change for each workflow or channel. This lets teams build on what works and expand to new use cases without starting over.
Presence brings together the components teams need to run agents in production: policies and standard operating procedures, guardrails, approved actions, simulations, evaluation tools, and a Codex-powered improvement process.
Together, these components help teams connect company systems, define how agents should behave, evaluate performance, enforce policies, and manage changes after launch.

Proven in leading enterprises
Built through years of deploying agents with customers, Presence is an ambitious product shaped by the demands of mission-critical environments. Every deployment generated insights that continuously inform OpenAI research and product development, creating compounding benefits for customers.
Presence powers OpenAI’s English-language phone support channel at 1-888-GPT‑0090, handling open-ended requests, verifying callers, using account context, and taking approved actions. Within weeks, it met or exceeded benchmarks we use to grade frontline human-support quality and now resolves 75% of inbound issues without human assistance. Working with our launch team, its Codex-powered improvement loop reduced human handoffs by 15 percentage points in just 10 days.
Leading enterprises are also building on the same proven foundation:
- BBVA is exploring AI-powered voice support for everyday banking needs in Mexico.
- SoftBank is testing natural Japanese-language customer conversations.
- IAG is exploring timely support during high-demand events such as severe weather.
1 of 3
Built for trust before, during, and after launch
Before a Presence deployment reaches users, teams can test it against common requests, edge cases, and higher-risk scenarios. Simulations and graders check whether it reached the right outcome, followed policy, used tools correctly, and escalated when appropriate. Guardrails can intervene when an interaction moves outside the company’s boundaries.

That work does not stop at launch. Production sessions, escalations, and quality signals can show teams where an agent is working well and where it needs attention as policies, products, and user behavior change. Codex using the Presence plugin investigates those signals and suggests updates. Teams can test each proposed change against the version in production, then approve a controlled rollout.
Presence is also designed to learn with you—improving as your business, customers, and employees evolve while the business stays in control. Built on OpenAI’s world-class research, it continues to advance as our models and capabilities improve. When a use case goes beyond what the product supports today, OpenAI Forward Deployed Engineers (FDEs) and partners can work with the customer to bring it into production.

Availability
OpenAI Presence is available to eligible enterprise customers as a deployed product through a limited general availability program. Deployments are led by OpenAI Forward Deployed Engineers and select global systems integrators. Presence is not yet available as a self-serve product.
As we introduce OpenAI Presence, we'll continue supporting voice customers with access to our frontier models through the OpenAI API.
To explore whether Presence is right for your organization, contact your OpenAI account team.
関連記事
今日のまとめ
AI日報で今日の重要ニュースをまとめ読み