OpenAI、開発者向けに音声信頼性とエージェント速度を向上させるAPIアップグレードを提供
本文の状態
日本語全文を表示中
詳細モードで約1分の本文を読めます。
同じ出来事の情報源
この情報源を基点に整理
The Decoder
OpenAIが新たな音声モデルと高速接続を導入し、開発者向けAPIの音声信頼性とエージェント処理速度を向上させました。
Source Article
元記事を日本語で読む
本文に関係しない購読案内、埋め込み通知、サイト内プロモーションは除いています。
OpenAI、開発者向けに音声信頼性とエージェント速度を向上させるAPIアップグレードを提供
OpenAIは開発者向けに2つのAPIアップデートを提供しました:リアルタイムAPI向けの新しいgpt-realtime-1.5モデルは、音声コマンドの信頼性を高めるために設計されています。内部テストでは、数字と文字の書き起こしが約10%向上し、論理的な音声タスクが5%向上、指示への追従が7%向上しました。音声モデルもバージョン1.5に更新されています。
Responses APIもWebSocketをサポートするようになりました。リクエストごとに完全なコンテキストを再送信する代わりに、新しいデータが入ってきたときだけ送信する永続的な接続を開きます。OpenAIによると、この変更により、多くのツール呼び出しを行う複雑なAIエージェントの速度が20〜40%向上します。
誇大広告なしのAIニュース – 人間によるキュレーション
THE DECODERの購読者として、広告なしの閲覧、週刊AIニュースレター、独占的な「AIレーダー」フロンティアレポート(年6回)、コメントへのアクセス、完全なアーカイブを入手できます。
原文を表示
OpenAI ships API upgrades targeting voice reliability and agent speed for developers
OpenAI has shipped two API updates for developers: the new gpt-realtime-1.5 model for the real-time API is designed to make voice commands more reliable. In internal testing, OpenAI saw roughly a ten percent improvement in transcribing numbers and letters, a five percent bump in logical audio tasks, and seven percent better instruction following. The audio model has also been updated to version 1.5.
The Responses API also now supports WebSockets. Instead of retransmitting the full context with every request, this opens a persistent connection that only sends new data as it comes in. According to OpenAI, the change speeds up complex AI agents with many tool calls by 20 to 40 percent.
AI News Without the Hype – Curated by Humans
As a THE DECODER subscriber, you get ad-free reading, our weekly AI newsletter, the exclusive "AI Radar" Frontier Report 6× per year, access to comments, and our complete archive.
関連記事
今日のまとめ
AIデイリーブリーフで今日の重要ニュースをまとめ読み