Qwen3:より深く思考し、より高速に動作する
Qwenチームは最新大規模言語モデル「Qwen3」を公開した。主力モデルと小型MoEモデルは、コーディングや数学で他トップモデルと互角の結果を示し、先行版を上回る性能を達成した。
QWEN CHAT GitHub Hugging Face ModelScope Kaggle DEMO DISCORD
紹介 本日、私たちは Qwen ファミリーの大規模言語モデルの最新作である Qwen3 のリリースを発表できることを嬉しく思います。当社のフラッグシップモデルである Qwen3-235B-A22B は、DeepSeek-R1、o1、o3-mini、Grok-3、Gemini-2.5-Pro などの他のトップティアモデルと比較して、コーディング、数学、一般能力などのベンチマーク評価において競争力のある結果を達成しています。さらに、小規模な MoE モデルである Qwen3-30B-A3B は、活性化パラメータが 10 倍の QwQ-32B を凌駕し、Qwen3-4B のような極小モデルでさえも Qwen2 のパフォーマンスに匹敵する能力を有しています。
原文を表示
QWEN CHAT GitHub Hugging Face ModelScope Kaggle DEMO DISCORD
Introduction Today, we are excited to announce the release of Qwen3, the latest addition to the Qwen family of large language models. Our flagship model, Qwen3-235B-A22B, achieves competitive results in benchmark evaluations of coding, math, general capabilities, etc., when compared to other top-tier models such as DeepSeek-R1, o1, o3-mini, Grok-3, and Gemini-2.5-Pro. Additionally, the small MoE model, Qwen3-30B-A3B, outcompetes QwQ-32B with 10 times of activated parameters, and even a tiny model like Qwen3-4B can rival the performance of Qwen2.
関連記事
News to Guide
ニュースの次に確認する
発表内容を、現在の料金や仕様と照らし合わせられる関連ガイドです。
今日のまとめ
AIデイリーブリーフで今日の重要ニュースをまとめ読み