AI エージェント向け検索 API ベンチ「Search Index」を公開
本文の状態
日本語全文を表示中
詳細モードで約3分の本文を読めます。
同じ出来事の情報源
この情報源を基点に整理
The Decoder
Artificial Analysis は AI エージェントの検索 API を品質、コスト、速度で評価する「Search Index」を公開し、Parallel や Exa などが上位にランクされた。
AI深層分析を開く2026年8月19日 04:18
AI深層分析
キーポイント
新ベンチマーク「Search Index」の発表
Artificial Analysis が AI エージェント向け検索 API の品質、コスト、速度を測定するベンチマーク「Search Index」を発表した。
厳格なテスト環境と主要競合の評価
GPT-5.6 Luna モデルを用いた標準化されたエージェント設定で 25 回の実行を行い、Parallel, Exa, Firecrawl, You.com などが評価対象となった。
コストと速度のトレードオフ分析
検索品質が高いほどトークン使用量が減り総コストが低下する一方、単発の応答速度が速くても品質が低ければ全体の処理時間は変わらないことが示された。
最適なバランスと公開方法
Parallel, Firecrawl, Parallel (turbo) がコストと性能のバランスで優れていると結論付けられ、他のプロバイダーもベンチマークへの参加を申請できる。
重要な引用
Artificial Analysis has released the "Search Index," a benchmark that measures how well search API providers work for AI agents across quality, cost, and speed.
Better search quality also lowers total costs. The model uses fewer tokens when it gets good results up front.
Raw speed per query doesn't always mean faster results overall.
編集コメントを表示
編集コメント
AI エージェントの実用化において、検索機能の選定は単なる速度競争ではなく、コスト効率と回答精度のバランスが鍵となる。このベンチマークは、開発者がデータに基づいて最適な API を選択するための重要な指針となるだろう。
Source Article
元記事を日本語で読む
本文に関係しない購読案内、埋め込み通知、サイト内プロモーションは除いています。
Artificial Analysis が「Search Index」というベンチマークを発表しました。これは、AI エージェントにとって検索 API プロバイダーが品質・コスト・速度の面でどの程度機能するかを測定するものです。
初期ラインナップには Parallel、Exa、Firecrawl、You.com、Tavily、Keenable、Brave が含まれています。各プロバイダーは、同じモデル(GPT-5.6 Luna)と標準化されたエージェント設定でテストされます。唯一変化する要素は検索プロバイダーです。エージェントは Artificial Analysis が提供するオープンソースフレームワークの Stirrup で動作し、各タスクあたり 25 回実行してウェブページを検索・取得します。
このインデックスは、3 つのベンチマークを等しく加重したものです。DeepSearchQA は 900 の研究質問を含み、それぞれ複数の検索クエリが必要です。BrowseComp サブセットでは、多段階のブラウジングが必要な 200 の難易度の高い事実がテストされます。AA-Omniscience は 6 つの知識ドメインにわたる 600 の質問をカバーします。比較基準として、モデルが独自に回答するツール非使用ベースラインも用意されています。

検索品質の向上は、総コストの削減にもつながります。モデルが初期段階で良好な結果を得られれば、使用するトークン数は減少します。Parallel Search(上級版)では、基本版と比較してトークン使用量が 40% 以上低下します。タスクあたりの検索コストは上昇しますが、トータルコストはむしろ低く抑えられます($0.084 vs $0.11)。
クエリごとの処理速度が速いからといって、必ずしも全体の結果取得時間が短いとは限りません。Parallel Search (turbo) はクエリあたりの応答時間(0.51 秒対 Basic の 1.03 秒)で最速を記録していますが、その品質スコア(67 対 73)が低いため、エージェントはより多くの試行回数を余儀なくされます。その結果、タスク全体の所要時間はほぼ同じになります。
Artificial Analysis によると、コストとパフォーマンスのバランスに優れているのは Parallel、Firecrawl、そして Parallel (turbo) です。他のプロバイダーも 参加を申請 してベンチマークに参加できます。なお、評価の詳細な手法は公開されています。
過剰な hype を排した AI ニュース – 人間が厳選
THE DECODER に登録すれば、広告なしでの閲覧、週刊の AI ニュレター、年 6 回の独占レポート「AI Radar」、アーカイブへの完全アクセス、そしてコメント欄の利用が可能になります。
原文を表示
Artificial Analysis has released the "Search Index," a benchmark that measures how well search API providers work for AI agents across quality, cost, and speed.
The initial lineup includes Parallel, Exa, Firecrawl, You.com, Tavily, Keenable, and Brave. Each one is tested with the same model (GPT-5.6 Luna) in a standardized agent setup. Only the search provider changes. The agent runs on Stirrup, an open-source framework from Artificial Analysis, and gets 25 runs per task to search and pull up web pages.
The index combines three equally weighted benchmarks. DeepSearchQA has 900 research questions, each requiring multiple search queries. A BrowseComp subset tests 200 hard-to-find facts that need multi-step browsing. AA-Omniscience covers 600 questions across six knowledge domains. A tool-free baseline, where the model answers on its own, provides the comparison point.

Better search quality also lowers total costs. The model uses fewer tokens when it gets good results up front. With Parallel Search (advanced), token use drops by over 40 percent compared to the Basic version. Per-task search costs go up, but total cost comes in lower ($0.084 vs. $0.11).
Raw speed per query doesn't always mean faster results overall. Parallel Search (turbo) clocks the shortest response time per query (0.51 seconds vs. 1.03 seconds for Basic), but its lower quality (67 vs. 73) forces the agent to run more passes. Total time per task winds up about the same.
Artificial Analysis says Parallel, Firecrawl, and Parallel (turbo) hit the best mix of cost and performance. Other providers can apply to join the benchmark. The full methodology is public.
AI News Without the Hype – Curated by Humans
Subscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive "AI Radar" frontier report six times a year, full archive access, and access to our comment section.
関連記事
News to Guide
ニュースの次に確認する
発表内容を、現在の料金や仕様と照らし合わせられる関連ガイドです。
今日のまとめ
AIデイリーブリーフで今日の重要ニュースをまとめ読み