Anthropicが「良しだが最高ではない」Claude Opus 4.7をリリース
本文の状態
日本語全文を表示中
詳細モードで約4分の本文を読めます。
同じ出来事の情報源
6媒体で確認
AI Business · InfoQ · AWS Machine Learning Blog · TechCrunch AI · Simon Willison Blog · The Decoder
各社の報じ方を比較 ↓AnthropicがClaude Opus 4.7をリリースし、モデルのドリフトや幻覚といった企業導入の主要な課題に対応することを目指している。
Source Article
元記事を日本語で読む
本文に関係しない購読案内、埋め込み通知、サイト内プロモーションは除いています。
Anthropic Releases Good but not Great Claude Opus 4.7
2分読取Anthropicは木曜日、最新モデルであるClaude Opus 4.7を発表し、限定公開されているサイバーセキュリティ向けモデルのClaude Mythosほど強力ではないと指摘した。
Anthropicによると、Opus 4.7はソフトウェアエンジニアリング(software engineering)の分野でOpus 4.6よりも進歩しており、初期ユーザーは監督なしである程度のコーディング作業をモデルに委ねられるという。Opus 4.7は長時間実行されるタスクを処理でき、指示により細かく従うことができる。また、画像の解釈にも優れており、スライドやドキュメントの作成においてより創造的になれる。
Opus 4.7はソフトウェアを簡単にハッキングできる能力で業界に衝撃を与えたClaude Mythosほど強力ではないものの、Anthropicはセキュリティ専門家はまだ脆弱性調査(vulnerability research)、ペネトレーションテスト(penetration testing)、レッドチーム演習(red teaming)などのサイバーセキュリティタスクに使用できると述べた。
「これは実際には、データベースをあるプラットフォームから別のプラットフォームへ移行するような、実際のビジネスタスクや困難なビジネスタスクにおいて、Opusが企業内でどのような位置を占めるかという話です」と、Futurum GroupのアナリストであるBradley Shimminは語った。
関連記事:The Real AI Shift Isn’t New Models. It’s Control.
Opus 4.7により、Anthropicはフロントティアモデル(frontier models)と連携する際に企業が直面する課題の解決を目指している。例えば、このベンダーはループ耐性(loop resistance)などの機能を含めており、これはモデルがユーザーの手動でリセットを要求する無限ループに陥るのを防ぎ、モデル内のハルシネーション(hallucinations)に対しても反発する働きをする。生成AIベンダーは、これらの機能によりモデルとの企業向け体験が大幅にスムーズになると主張している。
企業課題への対応Anthropicはまた、自社のモデルを単なるフロントティアモデルの1つではなく、企業をサポートするプラットフォームおよび基盤として構築することに意図的であると、Shimminは述べた。
モデルが基本コマンドやスキルをどのように管理するか、またはファイルシステムやメモリにどのようにアクセスするかなどの機能は、必ずしもモデルそのものの一部ではないが、パフォーマンスを決定する。Opus 4.7の自動モード(auto mode)などの新機能は、OpenAIやGoogleなどの競合他社のフロントティアモデルとは一線を画す、豊富なAPIと広範なメソドロジーを備え、プラットフォームのように動作するモデルを企業に提供することに注力していることを示していると、Shimminは補足した。
GartnerのアナリストであるArun Chandrasekaranによると、AnthropicがOpus 4.7で行った重要な改善点は、エージェント型推論能力(agentic reasoning capabilities)の向上にある。
「コーディングタスクに関するベンチマークが大幅に向上しており、これにはリファクタリング(refactoring)、[更新]されたレガシーコード、およびコードのドキュメント化などが含まれます」とChandrasekaranは述べた。「これらはすべてソフトウェア開発の分野であり、4.7モデルはこれらの領域で大幅に優れた結果を示しています。」
関連記事:OpenAI GPT-5.4-Cyber is More Open Than Claude Mythos
セキュリティの必要性しかし、Opus 4.7の機能向上にもかかわらず、このモデルはまだプレビュー段階のClaude Mythosほど強力ではない。
Shimminにとって、Anthropicは線を引きながら歩いているようだ。AIモデルが高度化するにつれ、ユーザーはエージェント型AI(agentic AI)を安全に使用する必要があると同時に、セキュリティを提供する必要があることを認識している。
「これは多くのアタックベクトル(attack vectors)を開く一方で、それらのアタックベクトルを解決するために使用できるという、多面的な状況に似ています」と彼は語った。
原文を表示
2 Min ReadAnthropic introduced its latest model, Claude Opus 4.7, on Thursday, noting that it is not as powerful as its limited-release cybersecurity model, Claude Mythos.Opus 4.7 is more advanced than Opus 4.6 in software engineering, with early users being able to hand off some coding work to the model without oversight, Anthropic said. Opus 4.7 can handle longer-running tasks and pay closer attention to instructions. It is also better at interpreting images and can be more creative when creating slides and documents. While Opus 4.7 is not as powerful as Claude Mythos -- which triggered a tech-industry scare because of its ability to easily hack software -- Anthropic said security professionals can still use it for cybersecurity tasks such as vulnerability research, penetration testing and red teaming. "This is a lot more about Opus finding its place within the enterprise for actual business tasks, difficult business tasks, like migrating a database from one platform to another,” said Bradley Shimmin, an analyst at Futurum Group. Related:The Real AI Shift Isn’t New Models. It’s Control.With Opus 4.7, Anthropic aims to address challenges enterprises face when working with frontier models. For example, the vendor included features such as loop resistance, which prevents the model from getting into a loop that requires the human user to start over, and even pushes back against hallucinations within the model. These features make enterprise experiences with models much smoother, the generative AI vendor claimed.Addressing Enterprise ProblemsAnthropic is also being more intentional about making its model more than just another frontier model, forming it into a platform and harness that supports enterprises, Shimmin said.Capabilities such as how a model manages basic commands or skills, or how it accesses the file system or memory, are not necessarily part of the model but determine how it performs. New features in Opus 4.7, such as auto mode, show the vendor is focusing on providing enterprises with a model that performs more like a platform, with a rich API and a wide range of methodologies that set it apart from frontier models from competitors such as OpenAI and Google, Shimmin added.A key improvement Anthropic made with Opus 4.7 is in agentic reasoning capabilities, said Arun Chandrasekaran, an analyst at Gartner."They have much better benchmarks in terms of coding tasks, which include everything from refactoring, [updating] legacy code and documentation of code,” Chandrasekaran said. "All of these are domains within software development where the 4.7 model shows much better results.”Related:OpenAI GPT-5.4-Cyber is More Open Than Claude MythosThe Need for SecurityHowever, even with the improved capabilities of Opus 4.7, the model is still not as powerful as Claude Mythos in preview.For Shimmin, it seems that Anthropic is walking the line, recognizing that as AI models become more advanced, users need agentic AI to be secured and also to provide security."It's like this multi-faceted sort of situation where it opens up a lot of attack vectors, but it can be used to solve a lot of attack vectors,” he said.About the AuthorNews Writer, AI BusinessEsther Shittu brings four years of expertise covering artificial intelligence technologies and industry trends. As co-host of the "Targeting AI" podcast, she talks to thought leaders and practitioners exploring critical AI developments. Previous to AI Business, she wrote for several publications including the New York Daily News, Bklyner and the Brooklyn Daily Eagle. When she's not diving deep into the world of AI, she spends her time on passion projects and raising her three daughters.
同じ出来事を6媒体で確認
同じ出来事を扱う別媒体の記事です。見出しと公開時刻を比較できます。
News to Guide
ニュースの次に確認する
発表内容を、現在の料金や仕様と照らし合わせられる関連ガイドです。
今日のまとめ
AIデイリーブリーフで今日の重要ニュースをまとめ読み