建設的回路増幅:標的サブネットワーク更新によるLLMの数学推論能力向上
本文の状態
日本語全文を表示中
詳細モードで約3分の本文を読めます。
同じ出来事の情報源
この情報源を基点に整理
Apple Machine Learning
LLM内部の特定タスクを担う「回路」と呼ばれる疎なサブネットワークを強化する手法を提案。標的的な更新により数学推論能力を向上させる研究。
Source Article
元記事を日本語で読む
本文に関係しない購読案内、埋め込み通知、サイト内プロモーションは除いています。
タイトル: 構成的回路増幅:ターゲットを絞ったサブネットワーク更新による大規模言語モデルの数学的推論能力の改善
構成的回路増幅:ターゲットを絞ったサブネットワーク更新による大規模言語モデルの数学的推論能力の改善
著者 Nikhil Prakash†, Donghao Ren, Dominik Moritz, Yannick Assogba
論文を見る
大規模言語モデル(LLM)の内部動作を調査したこれまでの研究では、特定のタスクを実行する責任を持つ、しばしば「回路」と呼ばれるスパースなサブネットワークが明らかにされてきた。さらに、ファインチューニングによるモデル性能の向上は、多くの場合、モデル内の既存の回路の強化から生じることが示されている。これらの知見を総合すると、このような回路に直接介入して、精密でタスクに特化した更新を行う可能性が示唆される。これらの発見に動機付けられ、我々は「構成的回路増幅」という新しい手法を提案する。この手法は、モデルの推論トレースから重要なトークンと、目的のタスクに関与するモデルコンポーネントを特定し、それらのコンポーネントのみを更新する。数学的推論に適用した結果、モデルコンポーネントのわずか1.59%を変更するだけで、複数のモデルにおいて精度を最大+11.4%向上させ、MMLU、TriviaQA、TruthfulQAで測定した他の能力への影響は最小限に抑えられた。これらの結果は、モデルコンポーネントのスパースなセットを選択的に更新することで、ターゲットを絞った能力を確実に強化できることを実証している。
† ノースイースタン大学
関連する文献と最新情報。
思考の幻想:問題の複雑性のレンズを通して推論モデルの強みと限界を理解する
2025年6月11日研究分野 音声・自然言語処理学会 NeurIPS
最近の最先端言語モデルの世代は、回答を提供する前に詳細な思考プロセスを生成する大規模推論モデル(LRM)を導入している。これらのモデルは推論ベンチマークで性能の向上を示すが、その基本的な能力、スケーリング特性、および限界については未だ十分に理解されていない。現在の評価は主に確立された数学およびコーディングのベンチマークに焦点を当てており、最終的な…
MUSCLE:互換性のあるLLM進化のためのモデル更新戦略
2024年10月23日研究分野 手法・アルゴリズム、研究分野 音声・自然言語処理学会 EMNLP
大規模言語モデル(LLM)は、通常、データやアーキテクチャの変更を通じて性能を向上させるために定期的に更新される。更新プロセスにおいて、開発者は多くの場合、全体的な性能指標の改善を優先し、以前のモデルバージョンとの互換性の維持にはあまり注意を払わない。あるモデルバージョンから次のバージョンへのインスタンスレベルの性能低下(インスタンス回帰)は、ユーザーのモデルに対するメンタルモデルを妨げる可能性がある…
機械学習における機会を発見する。
私たちの機械学習研究は、日々新たな領域を切り開いています。

原文を表示
Constructive Circuit Amplification: Improving Math Reasoning in LLMs via Targeted Sub-Network Updates
AuthorsNikhil Prakash†, Donghao Ren, Dominik Moritz, Yannick Assogba
View publication
Prior studies investigating the internal workings of LLMs have uncovered sparse subnetworks, often referred to as circuits, that are responsible for performing specific tasks. Additionally, it has been shown that model performance improvement through fine-tuning often results from the strengthening of existing circuits in the model. Taken together, these findings suggest the possibility of intervening directly on such circuits to make precise, task-targeted updates. Motivated by these findings, we propose a novel method called Constructive Circuit Amplification which identifies pivotal tokens from model reasoning traces as well as model components responsible for the desired task, and updates only those components. Applied to mathematical reasoning, it improves accuracy by up to +11.4% across multiple models while modifying as little as 1.59% of model components, with minimal impact on other abilities as measured by MMLU, TriviaQA, and TruthfulQA. These results demonstrate that targeted capabilities can be reliably enhanced by selectively updating a sparse set of model components.
† Northeastern University
Related readings and updates.
The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity
June 11, 2025research area Speech and Natural Language Processingconference NeurIPS
Recent generations of frontier language models have introduced Large Reasoning Models (LRMs) that generate detailed thinking processes before providing answers. While these models demonstrate improved performance on reasoning benchmarks, their fundamental capabilities, scaling properties, and limitations remain insufficiently understood. Current evaluations primarily focus on established mathematical and coding benchmarks, emphasizing final…
MUSCLE: A Model Update Strategy for Compatible LLM Evolution
October 23, 2024research area Methods and Algorithms, research area Speech and Natural Language Processingconference EMNLP
Large Language Models (LLMs) are regularly updated to enhance performance, typically through changes in data or architecture. Within the update process, developers often prioritize improving overall performance metrics, paying less attention to maintaining compatibility with earlier model versions. Instance-level degradation (instance regression) of performance from one model version to the next can interfere with a user’s mental model of the…
Discover opportunities in Machine Learning.
Our research in machine learning breaks new ground every day.

今日のまとめ
AIデイリーブリーフで今日の重要ニュースをまとめ読み