NVIDIA、LangChainのDeep Agentsでベンチマーク首位を達成
本文の状態
日本語全文を表示中
詳細モードで約6分の本文を読めます。
同じ出来事の情報源
この情報源を基点に整理
NVIDIA AI Blog
LangChain が最適化した Deep Agents ハーネスにより、NVIDIA Nemotron 3 Ultra はトップクローズドモデルより推論コストが約10分の1で、かつ最も高い精度とスループットを記録した。
AI深層分析を開く2026年8月4日 11:56
AI深層分析
キーポイント
コストと性能の両立達成
LangChain が最適化した Deep Agents ハーネスにより、NVIDIA Nemotron 3 Ultra はトップクローズドモデルより推論コストが約10分の1で、かつ最も高い精度とスループットを記録した。
環境最適化による革新
モデル自体の再学習やファインチューニングを行わず、システムプロンプトやツール記述などの周辺環境(ハーネス)を調整するだけで性能が向上したことが実証された。
オープンスタックの実装
NVIDIA NemoClaw と LangChain Deep Agents の組み合わせにより、企業がカスタマイズし、所有し、どこでも実行可能な完全なオープンスタックが提供されるようになった。
エンタープライズによるフルスタックの所有と柔軟な運用
オープンモデル、ハッチ、セキュアランタイムにより企業はエンドツーエンドでフルスタックを所有できる。これにより専門知識に基づいたカスタマイズが可能となり、自社のインフラやクラウド、ガバナンス下でどこでも実行可能となる。
エージェントの役割変化と高リスク業務への対応
AI アシスタントからコアシステム内で行動するエージェントへ移行することで、企業が得られる価値が変化する。この変化はよりリスクの高い業務を任せる際に特に重要となる。
重要な引用
"The way to build better agents is to keep improving the system around the model," said Harrison Chase, cofounder and CEO of LangChain.
"Our work with NVIDIA shows that enterprises can get strong performance from an open stack while keeping control over the agent systems they are building."
An open model, an open harness and an open secure runtime means enterprises own the full stack, end to end.
The shift from AI assistants that answer questions to agents that take action inside core systems changes what businesses get from their AI.
編集コメントを表示
編集コメント
モデルの再学習なしに性能が向上する「ハーネスエンジニアリング」のアプローチは、実務におけるコスト対効果と迅速な導入を可能にする画期的な知見である。企業にとって、クローズドモデルへの依存度を下げつつ自社のデータを完全に支配できるオープンスタックの実現は、今後の AI 戦略において重要な転換点となるだろう。
Source Article
元記事を日本語で読む
本文に関係しない購読案内、埋め込み通知、サイト内プロモーションは除いています。
NVIDIA Nemotron 3 Ultra は、世界で最も広く採用されている AI エージェントオーケストレーションプラットフォーム「LangChain」を活用することで、主要なクローズドモデルを上回る性能を、より低コストで実現しています。
LangChain は NVIDIA Nemotron 3 Ultra に最適化された Deep Agents ハーネスを調整し、オープンモデルの中で最高精度を達成しました。さらに、処理スループットも向上させ、実行あたりの推論コストは主要なクローズドモデルの 10 分の 1 に抑えています。
LangChain の Deep Agents ベンチマークでの評価では、Nemotron 3 Ultra は最高スコアを記録したクローズドモデルと同等のビジネスタスク処理能力を示しました。モデル自体の再学習は不要です。すべての性能向上は、モデルそのものではなく、モデルを取り巻く環境のエンジニアリングによるものです。
コストが 10 分の 1 で済むため、NVIDIA Nemotron 3 Ultra を活用するチームは評価を継続的に実行でき、より迅速に実験を行い、ビジネスの幅広い領域で専門的なエージェントを構築できます。
LangChain のエージェントエンジニアリングプラットフォームは、月間ダウンロード数が 2 億回を超えています。この Deep Agents ハーネスを NVIDIA Nemotron 3 Ultra に特化して調整することで、より多くのタスクを完了し、高速に動作する高性能なエージェントが可能になります。さらに、カスタマイズや所有、どこでも実行が可能な完全オープンスタックを企業に提供します。
「より優れたエージェントを構築するには、モデルを取り巻くシステム自体の改善が鍵となります」と語るのは、LangChain の共同創設者兼 CEO であるハリソン・チェイス氏です。「記憶機能やツールの活用、評価方法、そしてモデルの挙動は、チームがこれらを一体となって調整することで相乗効果を生みます。NVIDIA との連携を通じて、企業はオープンなスタックから強力なパフォーマンスを引き出しながらも、構築するエージェントシステムに対するコントロールを維持できることが示されました。」
Abridge、Amdocs、Box はそれぞれ専門的なエージェントを自社のプラットフォームやグローバルシステムに直接組み込んでいます。また、世界有数のシステムインテグレーターである EY は、LangChain Deep Agents 向けの NVIDIA NemoClaw ブループリンツを中心に、NVIDIA の実装能力を拡大しています。これにより、顧客は高価値なワークフロー全体で専門的なエージェントのカスタマイズ、評価、ガバナンスを支援されています。
NVIDIA の創設者兼 CEO ジェンソン・ファン氏は最近、チェイス氏と対談し、過去 6 ヶ月間でなぜ企業向け AI が飛躍的に有用になったのかについて議論しました。
微調整ではなくハネスのエンジニアリング
LangChain チームは、Nemotron 3 Ultra を公開されている Deep Agents ベンチマークスイートでテストし、実行トレースを分析してポイントが失われた箇所を特定しました。モデルの再学習を行うのではなく、チームはその周囲にあるハネス(システム枠組み)を調整したのです。具体的には、システムプロンプトやツールの説明、ミドルウェアなどを微調整しました。
Nemotron 3 Ultra と LangChain Deep Agents を活用するすべての開発者は、今日からこの成果を活用できます。調整済みのプロファイルは、LangChain を通じて直接利用可能です。
所有権を持つために設計されたオープンスタック
NVIDIA NemoClaw for LangChain Deep Agents は、独自の専門 AI(モデル・ツール・ランタイムからなるシステム)を構築する企業向けに、この成果をパッケージ化したオープンな参照設計図です。これは Nemotron 3 Ultra に最適化された「LangChain Deep Agents Code」と、エージェントのアクションを安全に実行するための NVIDIA OpenShell セキュアランタイムを組み合わせたものです。
モデル・ハネス(枠組み)・セキュアランタイムがすべてオープンであることは、企業がエンドツーエンドでフルスタックを完全に支配できることを意味します。自社の強みとなる専門知識に基づいてカスタマイズし、継続的に改善しながら、オンプレミスインフラでも自社クラウドでも、あるいは独自のガバナンス体制下でもどこでも運用可能です。
この違いは、エージェントがより重要な業務を引き受けるようになるときほど重要になります。質問に答える AI アシスタントから、コアシステム内で実際に行動を起こすエージェントへと移行することは、企業が AI から得られる価値そのものを変えます。
「NVIDIA NemoClaw for LangChain Deep Agents」と、最適化された Nemotron 3 Ultra モデルのプロファイルは現在利用可能です。開発者は、LangChain から直接最適化された Deep Agents ハネスを取得することもできますし、ゼロから専門的なエージェントを構築するための出発点として「NVIDIA NemoClaw for LangChain Deep Agents」の設計図を利用することもできます。
始め方
LangChain の開発者は、Baseten、Crusoe Cloud、DeepInfra、Fireworks、Nebius、Together AI の各プラットフォーム上で Nemotron 3 Ultra にアクセスでき、本格的な運用環境で最適化されたハネスを直接ホストされた形で利用できます。
EY は、このオープンなソフトウェアスタックを活用して、企業が今日から独自の専門エージェントの構築を開始できるよう支援します。
NVIDIA NemoClaw for LangChain Deep Agents や NVIDIA Nemotron に関する詳細はこちらをご覧ください。
アジェンティック AI、NVIDIA Nemotron など最新情報をお伝えするために、NVIDIA ニュースの購読やコミュニティへの参加、LinkedIn、Instagram、X、Facebook での NVIDIA AI のフォローをご検討ください。
自己学習型の動画チュートリアルやライブストリームもご活用いただけます。
原文を表示
NVIDIA Nemotron 3 Ultra is offering leading performance at lower cost than top closed models with the largest and most widely adopted AI agent orchestration platform.
LangChain tuned its Deep Agents harness for NVIDIA Nemotron 3 Ultra, achieving the highest accuracy among open models, while completing more tasks at higher throughput and running at 10x lower inference cost per run than leading closed models.
Measured against LangChain’s Deep Agents benchmark, Nemotron 3 Ultra also achieved business task parity with the highest-scoring closed models. No model retraining was required. Every gain came from engineering the environment around the model, not the model itself.
At a tenth of the cost, teams harnessing NVIDIA Nemotron 3 Ultra can run evaluations continuously, experiment faster and build specialized agents across more of their business.
LangChain’s agent engineering platform has more than 200 million monthly downloads. By tuning its Deep Agents harness specifically for NVIDIA Nemotron 3 Ultra, it allows for high-performing agents that complete more tasks, run faster and give enterprises a fully open stack they can customize, own and run anywhere.
“The way to build better agents is to keep improving the system around the model,” said Harrison Chase, cofounder and CEO of LangChain. “Memory, tool use, evaluation and model behavior compound when teams can tune them together. Our work with NVIDIA shows that enterprises can get strong performance from an open stack while keeping control over the agent systems they are building.”
Abridge, Amdocs and Box are embedding specialized agents directly into their platforms and global systems integrator EY is expanding its NVIDIA implementation capabilities around NVIDIA NemoClaw blueprints for LangChain Deep Agents, helping clients customize, evaluate and govern specialized agents across high-value workflows.
NVIDIA founder and CEO Jensen Huang recently sat down with Chase to discuss why the last six months have seen a leap in useful AI for enterprises.
Harness Engineering, Not Fine-Tuning
LangChain’s team ran Nemotron 3 Ultra against its public Deep Agents benchmark suite, then analyzed the deep agent’s execution traces to find exactly where it lost points. Instead of retraining the model, the team tuned the harness around it — adjusting system prompts, tool descriptions and middleware.
Every developer using LangChain Deep Agents with Nemotron 3 Ultra can put this to work today — the tuned profile is available directly through LangChain.
An Open Stack Built to Own
NVIDIA NemoClaw for LangChain Deep Agents is the open reference blueprint that packages this work for enterprises building their own specialized AI — systems of models, tools and runtime — tuned for their own workflows. It combines LangChain Deep Agents Code, tuned for Nemotron 3 Ultra, with the NVIDIA OpenShell secure runtime for executing agent actions safely.
An open model, an open harness and an open secure runtime means enterprises own the full stack, end to end. They can customize it around the expertise that sets their business apart, keep improving it and run it anywhere — their own infrastructure, their own cloud, their own governance.
That distinction matters more as agents take on higher-stakes work. The shift from AI assistants that answer questions to agents that take action inside core systems changes what businesses get from their AI.
NemoClaw for LangChain Deep Agents and the tuned Nemotron 3 Ultra model profile are available now. Developers can pull the tuned Deep Agents harness directly from LangChain, or use the NemoClaw for LangChain Deep Agents blueprint as a starting point for building specialized agents from scratch.
How to Get Started
LangChain developers can access Nemotron 3 Ultra on Baseten, Crusoe Cloud, DeepInfra, Fireworks, Nebius and Together AI platforms, giving them a direct, hosted path to the tuned harness in production.
EY can help enterprises start building their own specialized agents today, using this open software stack.
Learn more about NVIDIA NemoClaw for LangChain Deep Agents and NVIDIA Nemotron.
Stay up to date on agentic AI, NVIDIA Nemotron and more by subscribing to NVIDIA news, joining the community, and following NVIDIA AI on LinkedIn, Instagram, X and Facebook.
Explore self-paced video tutorials and livestreams.
News to Guide
ニュースの次に確認する
発表内容を、現在の料金や仕様と照らし合わせられる関連ガイドです。
今日のまとめ
AIデイリーブリーフで今日の重要ニュースをまとめ読み