字节跳动、最大10兆パラメータのAIモデルを開発中と報告
本文の状態
日本語全文を表示中
詳細モードで約1分の本文を読めます。
同じ出来事の情報源
この情報源を基点に整理
The Decoder
Bytedance は Financial Times の報道によると、10兆パラメータを持つ AI モデルを訓練中であり、これは既存の中国最大モデルである Moonshot の Kimi K3 を上回る規模となる。
Continue in AI NEW LAB
このニュースを、実務の判断につなげる
AI NEW LABで、試したことや先に確認したい条件を共有できます。まずはログインなしで読めます。
AI NEW LABで論点を見るAI深層分析を開く2026年8月7日 22:58
AI深層分析
キーポイント
超大規模モデルの開発計画
Bytedance は Financial Times の報道によると、10兆パラメータを持つ AI モデルを訓練中であり、これは既存の中国最大モデルである Moonshot の Kimi K3 を上回る規模となる。
独自学習アプローチの採用
内部関係者によると、Bytedance は過去 1 年以上にわたり他社のモデル出力を用いた蒸留(distillation)手法を避け、独自のデータとトレーニング方法で学習を進めている。
世界最高峰を目指す経営方針
創業者の張一鳴は Seed チームに対して長期的に世界をリードするモデル能力の実現を指示しており、xAI の Elon Musk も同規模のモデル開発を進めていると報じられている。
重要な引用
Bytedance is training an AI model with up to ten trillion parameters, according to the Financial Times.
One of the sources says Bytedance has avoided distillation, meaning training on outputs from other companies' models, for over a year.
Founder Zhang Yiming told the 2,000-person Seed team internally to aim for world-leading model capabilities over the long term.
編集コメントを表示
編集コメント
10 兆パラメータ規模への挑戦は、中国 AI エコシステムにおける技術的覇権争いの新たな局面を示している。蒸留手法を避けるという戦略は、長期的なモデルの独自性と品質向上に対する強い意欲の表れと言える。
Source Article
元記事を日本語で読む
本文に関係しない購読案内、埋め込み通知、サイト内プロモーションは除いています。
FTによると、ByteDanceは最大10兆パラメータを持つAIモデルの訓練を進めています。これは現在中国最大のモデルとされるMoonshotのKimi K3(約3.3兆パラメータ)の3倍に相当します。この規模は、業界推計で8兆パラメータ程度とされるAnthropicの最上位システム「Mythos 5」Mythos 5 と肩を並べる規模です。Anthropicは自社のパラメータ数を公表していません。
FTへの取材に応じた3人の関係者によると、このモデルは現在事前学習フェーズにあり、通常3〜6ヶ月を要する段階です。パラメータ数はモデルの記憶容量を決める重要な要素ですが、性能はデータの質や訓練手法にも大きく依存します。1人の情報提供者によれば、ByteDanceは過去1年以上にわたり、他社モデルの出力を用いた知識蒸留(distillation)には取り組んでいません。
創業者の張一鳴氏は、2,000人規模の「Seedチーム」に対し、長期的に世界トップクラスのモデル能力を追求するよう内部で指示を出しました。またElon Musk氏によると、xAIもColossus 2 クラスター上で6兆および10兆パラメータのGrokバリアントを訓練中です。
過剰な hype を排したAIニュース – 人間が厳選
広告なしで読める「THE DECODER」に購読してください。週刊AIニュースレター、年6回の独占レポート「AI Radar」、アーカイブ全記事へのアクセス、コメント欄の利用権などが付きます。
原文を表示
Bytedance is training an AI model with up to ten trillion parameters, according to the Financial Times. That's three times the size of Moonshot's Kimi K3, currently the largest Chinese model. It would put the TikTok parent company in the same ballpark as Anthropic's top system Mythos 5, which industry estimates place at around eight trillion parameters. Anthropic hasn't disclosed its own numbers.
Three insiders told the FT that the model is in pretraining, a phase that typically takes three to six months. Parameters determine how much a model can store, but performance also depends on data quality and training methods. One of the sources says Bytedance has avoided distillation, meaning training on outputs from other companies' models, for over a year.
Founder Zhang Yiming told the 2,000-person Seed team internally to aim for world-leading model capabilities over the long term. xAI is also training Grok variants with six and ten trillion parameters on its Colossus 2 cluster, according to Elon Musk.
AI News Without the Hype – Curated by Humans
Subscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive "AI Radar" frontier report six times a year, full archive access, and access to our comment section.
関連記事
今日のまとめ
AIデイリーブリーフで今日の重要ニュースをまとめ読み