Agnes AI が「Agnes 2.5 Pro Alpha」をリリース
本文の状態
日本語全文を表示中
詳細モードで約4分の本文を読めます。
同じ出来事の情報源
この情報源を基点に整理
Artificial Analysis
シンガポールのAIラボ、Agnes AI が推論モデル「Agnes 2.5 Pro Alpha」をリリースし、人工知能分析指数で39点を獲得しながらも低価格を実現した。
AI深層分析を開く2026年8月1日 13:57
AI深層分析
キーポイント
性能と価格のバランス
このモデルは推論能力において指数39を獲得し、同レベルの他社製品と比較して極めて低い価格帯で提供される。
コーディング能力の強み
コード生成に関する評価では指数58.8を記録し、類似価格帯のモデル群の中で上位に位置する。
学術的推論の実績
学術推論テストにおいて同レベルの他社製品と比較して競争力のあるスコアを示している。
コーディング性能の優位性
Agnes 2.5 Pro Alpha の Coding Index は 58.8 で、同価格帯の DeepSeek V4 Flash や GPT-5.4 mini を上回っている。
実務タスクにおける高いスコア
GDPval-AA v2 の評価では 1168点を獲得し、人間ベースライン(1,000点)を上回り GPT-5.4 mini と同等の性能を示した。
重要な引用
Agnes 2.5 Pro Alpha scores 39 on the Artificial Analysis Intelligence Index, placing it mid-pack among reasoning models we have tested and below the current proprietary frontier.
Coding is a relative strength for Agnes 2.5 Pro Alpha, scoring 58.8 on the Coding Index, near the top for its intelligence tier.
Agnes 2.5 Pro Alpha is strong in coding relative to its price tier: its Coding Index of 58.8 is above similarly priced models including DeepSeek V4 Flash (56.2), GPT-5.4 mini (56.1) and Qwen3.7 Plus (55.9).
On GDPval-AA v2, which rates performance on real-world work tasks against a human baseline of 1,000, Agnes 2.5 Pro Alpha scores 1168, above the baseline and level with GPT-5.4 mini (xhigh, 1169), ahead of Qwen3.7 Plus (943) but behind DeepSeek V4 Flash (max, 1189).
編集コメントを表示
編集コメント
Agnes AI は独自の基盤モデルを構築し、高機能でありながら低価格な推論モデルで市場に参入した。この発表は、AI利用におけるコスト構造の見直しを促す重要な事例となる。
Source Article
元記事を日本語で読む
本文に関係しない購読案内、埋め込み通知、サイト内プロモーションは除いています。
Agnes AI、低価格の推論モデル「Agnes 2.5 Pro Alpha」をリリース。Artificial Analysis Intelligence Index でスコア39、100万トークンあたり入力$0.45/出力$0.90
シンガポールに拠点を置く AI ラボ「Agnes AI」は、テキスト・画像・動画の全モダリティに対応する独自基盤モデルを自社開発し、300 万人以上のユーザーが利用する無料のオムニモーダル API を提供しています。今回発表された「Agnes 2.5 Pro Alpha」は、テキスト・画像・動画を入力として受け取り、テキストで回答を生成する推論モデルです。
Artificial Analysis Intelligence Index ではスコア39を獲得し、テスト済みの推論モデルの中では中位に位置します。ただし、現在の独自開発の最先端モデルには及びません。このモデルの最大の特徴は価格にあります。入力トークン 100 万あたり$0.45、出力トークン 100 万あたり$0.90という設定は、同程度の知能レベルを持つプロプライエタリモデルの中では極めて低価格です。
主な成果:
➤ Agnes 2.5 Pro Alpha は、Artificial Analysis Intelligence Index でスコア39を獲得し、リーダーボードへの初登場となりました。 これにより、現在の推論モデル群の中で中位に位置することが確認されました。
➤ コーディング能力は Agnes 2.5 Pro Alpha の相対的な強みです。Coding Index では 58.8 を記録し、同知能レベルの中ではトップクラスです。 これは、DeepSeek V4 Flash(56.2)、GPT-5.4 mini(56.1)、Qwen3.7 Plus(55.9)といった同価格帯のモデルを上回る性能で、Nex-N2-Pro(59.1)にわずかに及ばない程度です。
学術的な推論能力において、Agnes 2.5 Pro Alpha はその価格帯で競合他社と互角の性能を発揮しています。Humanity's Last Exam では 32%、GPQA Diamond では 88% のスコアを記録しました。
Humanity's Last Exam の結果は、同程度の価格設定を持つプロプライエタリモデルである GPT-5.4 mini (xhigh, 27%) や GPT-5.4 nano (xhigh, 26%) を上回っています。また、GPQA Diamond のスコアも、Intelligence Index で近い位置にある他社製品と同等の水準です。
1M トークンあたり入力・出力それぞれ$0.45/$0.90という価格設定は、同レベルの知能指数を持つモデルの中で 2 番目に安価です。キャッシュ入力・出力を 7:2:1 の比率でブレンドした場合、1M トークンあたりのコストは$0.18 に抑えられます。これは Artificial Analysis Intelligence Index で 38〜40 のスコアを持つモデル群の中で、DeepSeek V4 Flash (Reasoning, Max Effort) の$0.06 に次ぐ安さです。一方、Open weights の代替案である DeepSeek V4 Pro (Reasoning, Max Effort) や MiMo-V2.5-Pro も、より高い知能指数で同様の$0.18という価格を実現しています。
追加モデル詳細:
- コンテキストウィンドウ: 1M トークン
- 料金体系: 入力 1M トークンあたり$0.45 / 出力 1M トークンあたり$0.90 / キャッシュヒット 1M トークンあたり$0.0038
- 入力モダリティ: テキスト、画像、動画
- 利用方法: Agnes AI 公式 API

Agnes 2.5 Pro Alpha は、その価格帯においてコーディング能力に強みを持っています。Coding Index のスコアは 58.8 で、DeepSeek V4 Flash (56.2)、GPT-5.4 mini (56.1)、Qwen3.7 Plus (55.9) などと同程度の価格帯のモデルを上回っています。

GDPval-AA v2 は、実世界のタスクにおける性能を人間ベースライン(1,000 人)と比較して評価するベンチマークです。この基準で Agnes 2.5 Pro Alpha が得たスコアは 1168 で、人間ベースラインを上回り、GPT-5.4 mini (xhigh) の 1169 とほぼ同等の性能を示しました。Qwen3.7 Plus (943) を上回っていますが、DeepSeek V4 Flash (max) の 1189 には及びません。

Artificial Analysis Intelligence Index のタスクにおいて、Agnes 2.5 Pro Alpha が使用した出力トークンは 1 つあたり 22,000 トークンでした。そのうち推論に用いられたトークンは 16,000 トークンです。これは DeepSeek V4 Flash (max) の 45,000 トークンより 51% 少なく、GPT-5.4 mini (xhigh) の 78,000 トークンよりも 72% 少ない結果となりました。

Agnes 2.5 Pro Alpha の Artificial Analysis Intelligence Index に含まれる 10 の評価項目における詳細なパフォーマンスプロファイルです。GPQA Diamond や Humanity's Last Exam では特に高い結果を示す一方で、現在はまだ他モデルに劣っている分野も確認できます。

原文を表示
Agnes AI has launched Agnes 2.5 Pro Alpha, a low-priced reasoning model reaching 39 on the Artificial Analysis Intelligence Index at $0.45/$0.90 per 1M tokens
Agnes AI is a Singapore-based AI lab that trains its own full-modality foundation models in-house across text, image, and video, and offers them through a free omni-modal API that has passed 3 million users. Agnes 2.5 Pro Alpha is a text, image, and video input reasoning model with text output.
Agnes 2.5 Pro Alpha scores 39 on the Artificial Analysis Intelligence Index, placing it mid-pack among reasoning models we have tested and below the current proprietary frontier. Its defining characteristic is price: at $0.45 per 1M input tokens and $0.90 per 1M output tokens, Agnes 2.5 Pro Alpha is among the lowest-priced proprietary models at its intelligence level.
Key results:
➤ Agnes 2.5 Pro Alpha debuts at 39 on the Artificial Analysis Intelligence Index, its first appearance on our leaderboard. This places it mid-pack among current reasoning models.
➤ Coding is a relative strength for Agnes 2.5 Pro Alpha, scoring 58.8 on the Coding Index, near the top for its intelligence tier. This is above similarly priced models including DeepSeek V4 Flash (56.2), GPT-5.4 mini (56.1) and Qwen3.7 Plus (55.9), and sits just behind Nex-N2-Pro (59.1).
➤ On academic reasoning, Agnes 2.5 Pro Alpha is competitive for its tier, scoring 32% on Humanity's Last Exam and 88% on GPQA Diamond. Its Humanity's Last Exam result leads similarly priced proprietary models including GPT-5.4 mini (xhigh, 27%) and GPT-5.4 nano (xhigh, 26%), and its GPQA Diamond score is in line with peers near its Intelligence Index.
➤ At $0.45/$0.90 per 1M input/output tokens, Agnes 2.5 Pro Alpha is the second cheapest model at its intelligence level. Blended at 7:2:1 (cache-input-output) it costs $0.18 per 1M tokens, behind only DeepSeek V4 Flash (Reasoning, Max Effort) at $0.06 among models scoring 38-40 on the Artificial Analysis Intelligence Index. Open weights alternatives including DeepSeek V4 Pro (Reasoning, Max Effort) and MiMo-V2.5-Pro match that $0.18 price at higher intelligence.
Additional model details:
➤ Context window: 1M tokens.
➤ Pricing: $0.45 / $0.90 / $0.0038 per 1M input / output / cache hit tokens.
➤ Input modalities: Text, image, and video.
➤ Availability: Agnes AI first-party API.

Agnes 2.5 Pro Alpha is strong in coding relative to its price tier: its Coding Index of 58.8 is above similarly priced models including DeepSeek V4 Flash (56.2), GPT-5.4 mini (56.1) and Qwen3.7 Plus (55.9).

On GDPval-AA v2, which rates performance on real-world work tasks against a human baseline of 1,000, Agnes 2.5 Pro Alpha scores 1168, above the baseline and level with GPT-5.4 mini (xhigh, 1169), ahead of Qwen3.7 Plus (943) but behind DeepSeek V4 Flash (max, 1189).

Agnes 2.5 Pro Alpha used 22k output tokens per Artificial Analysis Intelligence Index task, 16k of them reasoning tokens, 51% fewer than DeepSeek V4 Flash (max, 45k) and 72% fewer than GPT-5.4 mini (xhigh, 78k).

Agnes 2.5 Pro Alpha's per-evaluation profile across the ten evaluations in the Artificial Analysis Intelligence Index, showing its stronger results on GPQA Diamond and Humanity's Last Exam alongside the areas where it currently trails.

AI算出
主要ニュースainew評価標準
AI モデルの新規リリースであり、独自ベンチマークや詳細な価格設定を含む具体的な数値情報が含まれているため新規性と関連性は高い。ただし、日本企業との直接的な関わりや日本語圏特有の情報は見当たらないため、日本の関連性は低めとした。
6つの評価軸を見る
- AI関連度
- 100
- 情報源の信頼性
- 25
- 新規性
- 25
- 調べる価値
- 25
- 重複の少なさ
- 100
- 日本での有用性
- 25
関連記事
News to Guide
ニュースの次に確認する
発表内容を、現在の料金や仕様と照らし合わせられる関連ガイドです。
今日のまとめ
AIデイリーブリーフで今日の重要ニュースをまとめ読み