Grok 4.6
Grok 4.6
コメントの核心
Grok 4.6 の利用体験と性能評価
意見が分かれるGrok 4.5 および 4.6 は、他の主要モデルと比較して高速で冗長性がない点で好意的に評価されている。セキュリティレビューにおいて攻撃経路を特定する能力が高く、UI も改善されたとの報告がある。一方で、生成される計画が散漫で自己矛盾を起こしやすく、指示に従う能力に欠けると感じるユーザーも存在する。
代表コメント(原文)
“In terms of using experience, I found Grok 4.5 to be way more pleasant to use than GPT 5.6 Sol and Claude 4.8/5. It just gets to the point, and is…”
他社モデルとの競争とベンチマークの信頼性
意見が分かれるGrok の登場が業界に健全な競争をもたらすという見方がある。しかし、短期間で同様の性能向上が見られる点やベンチマーク結果の信憑性については懐疑的な意見も根強い。Gemini が真の競争相手となる可能性についても言及されているが、現状ではまだ時間がかかるという評価もある。
代表コメント(原文)
“Anyone else find it weird how within 2 months of Fable releasing all the major labs suddenly had Fable-level models? Trying to think of explanations:…”
全論点と英語の原コメントを見る
議論の全体像
Grok 4.6 の発表を受け、開発者間で利用体験や性能評価が分かれている。一部は GPT や Claude よりも高速で要領が良いと支持し、セキュリティレビューでの成果を称賛する声がある一方、計画の整合性や指示従順性に疑問を呈する意見も見られる。また、他社モデルとの比較やベンチマークの信頼性についても議論が交わされた。
Grok 4.6 の利用体験と性能評価
Grok 4.5 および 4.6 は、他の主要モデルと比較して高速で冗長性がない点で好意的に評価されている。セキュリティレビューにおいて攻撃経路を特定する能力が高く、UI も改善されたとの報告がある。一方で、生成される計画が散漫で自己矛盾を起こしやすく、指示に従う能力に欠けると感じるユーザーも存在する。
“In terms of using experience, I found Grok 4.5 to be way more pleasant to use than GPT 5.6 Sol and Claude 4.8/5. It just gets to the point, and is super fast and concise, no yapping. That's how AI agents should be imo. None of the weird "Claude ipsum" jargon…”
“I had a few interactions with it through cursor and first impression is: I'm underwhelmed. The plans it produces are all over the place and hard to follow. They have a "rambly" feel to it. Worse, they start becoming self contradictory after a few rounds of…”