OpenAI、モデルの「ゴブリン」発言禁止問題について言及
本文の状態
日本語全文を表示中
詳細モードで約3分の本文を読めます。
OpenAI は、自社のコーディングモデルがゴブリンや妖精などの架空の生物を話題にしないよう指示された事実に言及し、これをモデルが発達させた奇妙な習慣であると説明した。
Continue in AI NEW LAB
このニュースを、実務の判断につなげる
AI NEW LABで、試したことや先に確認したい条件を共有できます。まずはログインなしで読めます。
AI NEW LABで論点を見るSource Article
元記事を日本語で読む
本文に関係しない購読案内、埋め込み通知、サイト内プロモーションは除いています。
GPT-5.1 の「おたく」人格のリリースをきっかけに、ゴブリンやグレムリンへの言及が急増し、その後他のモデルにも広がりました。
GPT-5.1 の「おたく」人格のリリースをきっかけに、ゴブリンやグレムリンへの言及が急増し、その後他のモデルにも広がりました。
by Emma Roth
2026 年 4 月 30 日 午後 10:42 GMT+9


画像: The Verge
Emma Roth
ストリーミング戦争、消費者向けテクノロジー、暗号資産(クリプト)、ソーシャルメディアなどを担当するニュースライターです。以前は MUO でライターおよび編集者を務めていました。
OpenAI は自社の「ゴブリン問題」について語り始めました。Wired による報道 [1] が、OpenAI のコーディングモデルに対して「ゴブリン、グレムリン、アライグマ、トロール、オーガ、ハト、その他の動物や生物については決して話してはならない」という指示が組み込まれていることを明らかにした直後、同 AI スタートアップは自社のウェブサイトに解説を掲載し、これらの生物への言及はモデルの学習結果として生じた「奇妙な習慣」であると説明しました。
[1] https://www.wired.com/story/openai-really-wants-codex-to-shut-up-about-goblins/
[2] https://openai.com/index/where-the-goblins-came-from/
ブログ記事で概説されている通り、OpenAI は GPT-5.1 モデルから [「Nerdy」の人格オプションを使用する際] に、ゴブリンや他の生物を参照する比喩に気づき始めました。OpenAI によると、この問題は後続のモデルリリースでも悪化し続け、最終的に強化学習(Reinforcement Learning)が「Nerdy」の人格を持つモデルにおいて、奇抜な比喩に対して報酬を与えていることが判明しました。その結果、新しいモデルはそれを学習対象としてしまいました。
**この報酬は「Nerdy」条件でのみ適用されましたが、強化学習では、学習された行動が生成された条件にきれいに限定されて維持されるとは限りません。一度スタイルの癖(Style Tic)が報酬で強化されると、後のトレーニングによってそれが他の場所へ広まったり、さらに強化されたりする可能性があります。特に、その出力が教師あり微調整(Supervised Fine-tuning)や選好データ(Preference Data)で再利用される場合、その傾向は顕著になります。
OpenAI が「Nerdy」という人格を3月に廃止した後も、ゴブリンやグレムリンへの言及は完全に消えたわけではありませんでした。GPT-5.5 が Codex コーディングツール内に登場した際も、同社は「根本原因」を発見する前にモデルのトレーニングを開始していたためです。その結果、会社は Codex に対して、これらの神話上の生物について言及しないよう非常に具体的な指示を出す必要がありました。しかし、AI にゴブリンを少し混ぜてコードを書きたいとお考えの場合は、OpenAI は その指示を元に戻す方法 を共有しています。
この記事のトピックや著者をフォローして、パーソナライズされたホームページフィードで類似の記事をもっとご覧いただき、メールでの更新通知を受け取ってください。
- エマ・ロス
-
-
The Verge Daily
最も重要なニュースを毎日お届けする無料ダイジェスト。
メールアドレス(必須)
原文を表示
References to goblins and gremlins spiked with the release of GPT-5.1’s ‘Nerdy’ personality, and then spread to other models.
References to goblins and gremlins spiked with the release of GPT-5.1’s ‘Nerdy’ personality, and then spread to other models.
by Emma Roth
Apr 30, 2026, 10:42 PM GMT+9


Image: The Verge
Emma Roth
is a news writer who covers the streaming wars, consumer tech, crypto, social media, and much more. Previously, she was a writer and editor at MUO.
OpenAI is opening up about its goblin problem. After a report from Wired revealed instructions to OpenAI’s coding model to “never talk about goblins, gremlins, raccoons, trolls, ogres, pigeons, or other animals or creatures,” the AI startup published an explanation on its website, calling references to the creatures a “strange habit” its models developed as a result of their training.
As outlined in the blog post, OpenAI began noticing metaphors referencing goblins and other creatures starting with its GPT-5.1 model — specifically when using the “Nerdy” personality option. OpenAI says the problem continued to worsen with subsequent model releases, until it found that its reinforcement training rewarded the quirky metaphors with the Nerdy personality, which newer models were training on.
The rewards were applied only in the Nerdy condition, but reinforcement learning does not guarantee that learned behaviors stay neatly scoped to the condition that produced them. Once a style tic is rewarded, later training can spread or reinforce it elsewhere, especially if those outputs are reused in supervised fine-tuning or preference data.
Though references to goblins and gremlins dropped off after OpenAI discontinued the Nerdy personality in March, they didn’t disappear completely with GPT-5.5 inside its Codex coding tool, as OpenAI started training the model before finding the “root cause.” The company had to give Codex very specific instructions not to talk about the mythological creatures as a result. But if you’d prefer to have your AI code with some goblin sprinkled in, OpenAI has shared a way to reverse its instructions.
Follow topics and authors from this story to see more like this in your personalized homepage feed and to receive email updates.
- Emma Roth
-
-
-
The Verge Daily
A free daily digest of the news that matters most.
Email (required)
同じ出来事を3媒体で確認
同じ出来事を扱う別媒体の記事です。見出しと公開時刻を比較できます。
関連記事
今日のまとめ
AIデイリーブリーフで今日の重要ニュースをまとめ読み