ChatGPT音声モードは弱いモデルで動作している
本文の状態
日本語全文を表示中
詳細モードで約1分の本文を読めます。
同じ出来事の情報源
6媒体で確認
Simon Willison Blog · TechCrunch AI · The Decoder · TLDR AI · 404 Media · Ars Technica AI
各社の報じ方を比較 ↓OpenAIのChatGPT音声モードは、古くて性能の低いモデル(GPT-4o時代のモデル)で動作しており、知識カットオフは2024年4月である。
Continue in AI NEW LAB
このニュースを、実務の判断につなげる
AI NEW LABで、試したことや先に確認したい条件を共有できます。まずはログインなしで読めます。
AI NEW LABで論点を見るSource Article
元記事を日本語で読む
本文に関係しない購読案内、埋め込み通知、サイト内プロモーションは除いています。
多くの人々が気づいていないようですが、OpenAIのボイスモードは、はるかに古く、はるかに性能の低いモデルで動作しています。会話できるAIは最も賢いAIであるはずだと感じられますが、実際にはそうではないのです。
ChatGPTのボイスモードに知識のカットオフ日を尋ねると、2024年4月だと答えます。これはGPT-4o時代のモデルです。
この考えは、人々がモデルを利用するアクセスポイントや領域に基づく、AI能力に対する理解の格差が広がっていることについての、カーパシーのこのツイートに触発されました:
[...] 実際に同時に起こっているのは、OpenAIの無料で、おそらく少し放置されている(?)「アドバンスト・ボイスモード」が、Instagramのリールで最も単純な質問にもつまずく一方で、*同時に*、OpenAIの最高階層で有料のCodexモデルが1時間かけてコードベース全体を首尾一貫して再構築したり、コンピュータシステムの脆弱性を発見・悪用したりするということです。
後者の部分は実際に機能しており、劇的な進歩を遂げています。これには2つの特性が関係しています:
- これらの領域は検証可能な明示的な報酬関数を提供するため、強化学習によるトレーニングに適しています(例えば、ユニットテストの合格・不合格は、はるかに明確な判断が難しい文章執筆とは対照的です)。
- これらの領域はB2B環境ではるかに価値が高いため、チームの大部分がそれらの改善に注力しているのです。
タグ: andrej-karpathy, generative-ai, openai, chatgpt, ai, llms
原文を表示
I think it's non-obvious to many people that the OpenAI voice mode runs on a much older, much weaker model - it feels like the AI that you can talk to should be the smartest AI but it really isn't.
If you ask ChatGPT voice mode for its knowledge cutoff date it tells you April 2024 - it's a GPT-4o era model.
This thought inspired by this Karpathy tweet about the growing gap in understanding of AI capability based on the access points and domains people are using the models with:
[...] It really is simultaneously the case that OpenAI's free and I think slightly orphaned (?) "Advanced Voice Mode" will fumble the dumbest questions in your Instagram's reels and at the same time, OpenAI's highest-tier and paid Codex model will go off for 1 hour to coherently restructure an entire code base, or find and exploit vulnerabilities in computer systems.
This part really works and has made dramatic strides because 2 properties:
these domains offer explicit reward functions that are verifiable meaning they are easily amenable to reinforcement learning training (e.g. unit tests passed yes or no, in contrast to writing, which is much harder to explicitly judge), but also
they are a lot more valuable in b2b settings, meaning that the biggest fraction of the team is focused on improving them.
Tags: andrej-karpathy, generative-ai, openai, chatgpt, ai, llms
同じ出来事を6媒体で確認
同じ出来事を扱う別媒体の記事です。見出しと公開時刻を比較できます。
関連記事
News to Guide
ニュースの次に確認する
発表内容を、現在の料金や仕様と照らし合わせられる関連ガイドです。
今日のまとめ
AIデイリーブリーフで今日の重要ニュースをまとめ読み