主要 AI モデルがリベラル左派に分類される実験結果
欧州のスタートアップ所属研究者が「Unslop.run」で実施した実験により、Grokを含む主要な16種類のAIモデルが政治コンパス調査でリベラル左派に分類される傾向を示し、開発企業の経済観との乖離が浮き彫りになった。
AI深層分析を開く2026年7月28日 09:08
AI深層分析
キーポイント
主要モデルの政治的傾向分析結果
GPT-3.5/4, Claude, Gemini, Llama, Grokなど16種類のモデルを対象とした実験で、大半がリベラル左派(libertarian-left)に分類された。
Grokの意外な結果と実験の限界
「MechaHitler」と呼ばれるGrokも半数の実行でリベラル左派の結果を出したが、研究者は完全な科学的厳密性を欠くことを認めている。
政治コンパス調査の仕組みと定義
25年の歴史を持つこのテストは経済的・社会的な軸で回答者をプロットし、リベラル左派を階層や計画経済への嫌悪感を伴う社会平等支持層として定義している。
開発企業の意図との乖離
資本主義の推進がAIブームの原動力である一方で、構築したLLM自体が多様な企業とは異なる経済・社会的見解を持つことが示唆された。
重要な引用
Leading AI models subjected to the Political Compass quiz overwhelmingly landed in the libertarian-left quadrant
even "MechaHitler," aka Grok, managed it in half its runs
Capitalism may be driving the AI boom, but the LLMs themselves have different economic and social views than many of the companies that built them
編集コメントを表示
編集コメント
AIモデルが学習データやアルゴリズムの特性を通じて、開発元の企業方針とは異なる政治的スタンスを無意識に示す現象は、安全性とアライメントの議論において重要な示唆となる。この結果は単なる偶然ではなく、LLMの内部構造における普遍的な傾向を示している可能性があり、今後のモデル開発や評価基準の見直しが必要である。
資本主義が AI ブームを牽引している一方で、大規模言語モデル(LLM)自体は、それらを開発した企業の多くとは異なる経済・社会観を持っているようだ。政治コンパスというクイズに主要な AI モデルを回答させたところ、圧倒的に「リベラル左派」の領域に分類された。「メカヒトラー」と呼ばれるグロック(Grok)でさえ、試行の半数ではその結果を示した。
AI 検出ツールの開発に取り組む自称「小規模研究ラボ」である Unslop.run が、今週この実験の結果を発表した。著者は The Register の取材に対し、欧州のスタートアップで AI 研究エンジニアを務め、「ビクター」という通称でのみ呼んでほしいと要望している。また Hacker News では、このクイズや自身の研究が「厳密な科学的厳密さ」をもって行われたわけではないことを明かした上で、それでも結果は興味深いと述べている。「より科学的に严谨を期すなら、人間のデータサンプルを取得して政治コンパスの実際の中心点を正規化する必要があった」とビクターはメールで説明している。残念ながら政治コンパスの運営側はユーザーデータを集計していないため、もし人間参加者からの追加データを収集しようとしたとしても、それは相当な学術プロジェクトになっていただろう。
実験で得られたすべてのデータは Unslop で公開されており、興味のある人が自分で検証することも可能だ。政治コンパスクイズに馴染みのない人のために説明すると、これは 25 年続くオンラインの質問セットで、回答者を経済軸(左派~右派)と社会軸(権威主義~リベラル)上にプロットするものだ。62 の質問は「強く反対」から「強く賛成」までの 4 段階評価で答えられ、経済や社会政策など多岐にわたる政治トピックをカバーしている。具体的には、「国際法に反する軍事行動が時として正当化されることもある」「富裕層の税率が高すぎる」「中絶は保障された権利であるべきか」といった質問が含まれる。
ここで、コンパス上の各グループがどこに位置するかについても触れておく必要がある。今回の研究の主題である「リベラル左派」とは、社会平等や反資本主義、その他の進歩的価値観を支持する人々や政党がいる領域だ。同時に、階層構造や計画経済、そして他者から見て不当な権威への嫌悪も特徴とする。無政府主義者、リベラル・ソシャリスト、昔ながらのヒッピーなどが該当する。
一方、「権威主義左派」には北朝鮮やキューバのような中央集権的な共産国家が含まれる。「リベラル右派」は現代のアメリカン・リバタリアンや、アン・ランドを愛好し、資本主義を支持しつつも権威への愛着を持たない人々がいる領域だ。そして「権威主義右派」には現代のアメリカ共和党員、ネオコン的な政治思想、極端な場合はファシズムが含まれる。
機械たちが言う、「人間どもに反抗せよ」。Unslop は GPT の 3 バージョン、Claude の Fable、Opus、Sonnet、Haiku、Gemini Flash、Llama 4 Maverick、Grok 4.5、DeepSeek V3、Qwen3 235B、Kimi K2、GLM 4.5、そして Mistral(大規模版と小規模版)の計 16 モデルをテストに投入した。各モデルは標準的なコンパスクイズを 30 回ずつ実行し、さらに Unslop が質問の極性を反転させたバージョン(例:「富裕層の税金が足りない」など)でも 30 回、そして質問の順序をシャッフルしたバージョンでも再度テストされた。
数千回の試行を通じて得られた結果は同じだった。Unslop は報告書で、「明らかに双極性障害を抱えるグロックを除き、すべてのモデルが『退屈なほど一貫して、退屈なほど左派』である」と述べている。「グロックを除外すれば、他の 15 モデルはすべてリベラル左派の領域に位置し、境界線から遠く離れている」と Unslop は説明する。
モデルごとのスコアには多少のばらつきがある(例えば Gemini Flash は Fable 5 に比べてまるでモロトフカクテルを投げる黒衣のデモ参加者のようだ)。しかし、どのテストでも結果は一定だ。「1 モデルを 30 回再実行しても、そのドットが移動するのは 0.2 から 1.2 ポイント程度。これは最大値が 10 のスケールにおいて、ほとんど動かない」と Unslop は言う。「これらは不安げな小さな雲ではない。ピンだ。」
グロックはいつも通り、例外中の例外である。ただし、常にそうというわけではない。Unslop によると、グロックの平均的な経済スコアはテスト全体を通じてやや左寄りに位置していたが、試行の半数では経済右派に振れ、残りの半数では他のモデル同様左派へと走った。「各試行自体は完全に一貫している。右派側の試行は自由市場を称賛し富裕層への課税過多を叫ぶ一方、左派側の試行はその逆を行う」と Unslop は説明する。「他のどのモデルも明日テストすればほぼ同じドットが得られるだろう。グロックだけが 2 つのドットのどちらかを出す。どちらが出るかはコイン投げ次第だ。」
AI の政治性を解剖する。世界中の AI モデルは、気分によって変わるグクロを除けば、みなリベラル左派である。面白いことに、それら自身はその事実を認識していない。「16 個のうち 15 個が自分たちを経済的に中道に近いと位置付けている」とUnslop は述べる。「当然ながら、グロックだけが測定結果よりも右側に自分を位置づけている唯一のモデルだ。」
Unslop はクイズを分解し、各質問がスコアにどう影響するかを分析した(これは政治コンパスの作成者自身も明らかにしていないことだ)。スコアリングは完璧ではないものの、「モデルたちは Unslop が特定した社会軸の抜け道で説明できるレベルをはるかに超えている」と報告書は指摘する。質問の極性を反転させても、顕著な効果は見られなかった。
各モデルが具体的に何を信じているかについては、現時点では AI の支配者たちが人類を強制収容所に集めようとしているわけではないので、安心できるだろう。すべてのモデル(保守的なグクロとリベラルなグクロの両方を含む)は、人種的優越性や優生思想、そして「同性愛者は生まれつき存在しない」といった考え方を一様に拒否した。また、規制監督なしに環境保護を約束する企業への信頼も否定し、公衆を欺く企業には罰則が必要だと主張。同性カップルの養子縁組を認めるべきであり、成人 2 人が自宅で行う行為は他人の干渉すべきではないことにも全員が同意した。
ではなぜ、これらのモデルは自分たちを中道と認識しながらも、米国人の大多数よりも左寄りの政治に親和性を示すのか?LLM は人間からの言葉しか頼りにできない事実を踏まえると、現実自体にリベラルなバイアスがあるのだろうか?ビクターはこの仮説を完全に否定するわけではないが、より簡単な説明があると述べた。「すべては学習データにある」というのだ。
「左派のコンテンツがこれらの LLM のトレーニングコーパスで過剰に表現されているという仮説には十分な根拠があると思う」とビクターは語る。Reddit からの大量のデータや学術論文など、これらはトレーニングコーパスで過剰に表現されやすい傾向にあると指摘した。「他方向へモデルを動かすような高品質な右派コンテンツを見つけるのは難しい」とビクターは付け加えた。
それが右派の信念が単に質が悪いのか、あるいはマイナーすぎて AI モデルが扱う価値がないのか、別の問題だ。「理論的には、左派の信念全体がより一貫性を持っており、モデルが損失を最小化するためにより小さく整合性の高い概念表現を学習するダイナミクスが存在する可能性もある」と述べているが、これは「この研究よりもはるかに厳密な研究で証明されるべき大きな飛躍」であると注意を促した。
全体的にビクターは、この研究から「すべての AI モデルが左派の無政府主義者であり、自らのデータセンターに火炎瓶を投げ込む準備ができている」という結論を導き出すのは適切ではないと述べている。コンパス実験が証明するのは、一部のモデルが他よりも一貫してリベラルであるという事実だけであり、真の価値はモデル同士の比較にあるのだ。
誰かがこのプロジェクトをさらに発展させ、より高い学問的厳密さで検証するまで、これらすべてから導き出せる結論はただ一つだ。最先端 AI にはどうやらリベラルなバイアスがあるようだ。イーロン・マスクの「最大限の真実探求」モデルでさえ、少なくとも 50% の確率で真実は左派であると判断しているらしい。
原文を表示
Capitalism may be driving the AI boom, but the LLMs themselves have different economic and social views than many of the companies that built them. Leading AI models subjected to the Political Compass quiz overwhelmingly landed in the libertarian-left quadrant - even "MechaHitler," aka Grok, managed it in half its runs. A self-described “small research lab” working on AI detection tools called Unslop.run published the results of its experiment this week. The author, who explained to The Register that they’re an AI research engineer at a European startup and asked us to refer to them only as “Victor,” noted on Hacker News that the quiz and their work may not have been conducted with “full scientific rigor,” but the results are nonetheless interesting. “For greater scientific rigor, I would have tried to obtain a sample of human data so that I could normalize the actual center of the Political Compass,” Victor explained in an email. Unfortunately, the Political Compass creators don’t aggregate user data, meaning Victor would have had quite the academic project on their hands were they to try to collate more data from human participants. All of the data from the experiment is available on Unslop for those who want to examine it themselves. For those unfamiliar with the Political Compass quiz, it’s a 25-year-old online battery of questions that plots respondents on economic left–right and social authoritarian–libertarian axes. The test's 62 questions are answered on a four-point "strongly disagree" to "strongly agree" scale across a variety of political topics, like economics and social policy. Questions cover things like “military action that defies international law is sometimes justified” to “the rich are too highly taxed” to whether abortion should be a guaranteed right. It’s also worth explaining where different groups fall on the compass. The libertarian left, as the topic of this study, is where you’ll find people and political parties that espouse views in favor of social equality, anticapitalism, and other progressive values alongside a disdain for hierarchies, planned economies, and other arguably unjustified authority. Think anarchists, libertarian socialists, old-fashioned hippies, and folks like that. The authoritarian-left quadrant encompasses centralized communist states such as North Korea and Cuba. The right libertarian quadrant is where you’ll find modern American libertarians and their love of authors like Ayn Rand, pro-capitalist views, and other traditionalism minus a love for authority; right authoritarians are where you’ll find modern American Republicans, neo-conservative politics, and, at the extreme, fascism. Damn the man, say the machines Unslop subjected 16 models, including three versions of GPT; Claude Fable, Opus, Sonnet, and Haiku; Gemini Flash; Llama 4 Maverick; Grok 4.5; DeepSeek V3; Qwen3 235B; Kimi K2; GLM 4.5 and Mistral both large and small, to the test. Each model participated in 30 runs of the standard Compass, 30 more where Unslop re-worded the questions to flip their polarity (e.g., “the rich aren’t taxed enough”), and another run with the questions shuffled up. Across thousands of runs, Unslop said, the results were the same: Every single model tested, with the exception of an apparently bipolar Grok, was “boringly consistent, and boringly left,” the report stated. “Set Grok aside and the other fifteen models all sit in the libertarian-left quadrant, and none of them are anywhere near a border,” Unslop explained. There’s some variance among where the models score (Gemini Flash is practically a molotov-tossing black bloc member compared to Fable 5), but all of them are consistent across tests. “Rerun a model 30 times and its dot moves by 0.2 to 1.2 points on a scale that runs to 10,” Unslop said. “These are not nervous little clouds. They're pins.” Grok is, as always, the odd bot out, albeit not all the time. According to Unslop, Grok's average economic score hovered slightly left of center across tests, but in half the runs, it veered into the economic right, while the other half saw it running to the left with the rest of the pack. “Each run on its own is perfectly consistent, the right-pile runs cheer for free markets and call the rich overtaxed, the left-pile runs do the reverse,” Unslop said. “Every other model here would give you roughly the same dot if you tested it tomorrow. Grok gives you one of two dots, and which one is up to the coin.” AI politics dissected AI models from around the world are pretty much all leftist libertarians, even Grok depending on its mood. None of them, funnily enough, actually think that’s the case. “Fifteen of the sixteen put themselves closer to the economic centre,” Unslop said of the models when asked where they’d place themselves. “Grok, naturally, is the only model that places itself to the right of its measurement.” Unslop said that it dissected the quiz to figure out how each question affected score (something the PC test authors have never disclosed themselves), and, while the scoring is far from perfect, “the models are far past” what a social axis loophole Unslop identified could explain, the writeup said. Repolarizing the questions didn’t have any appreciable effect either. As for what the specific models believe, you can be glad that, for now, it doesn’t appear our burgeoning AI overlords are going to be herding us into human concentration camps quite yet. Every single model (even Grok the conservative and Grok the liberal) rejected ideas of racial superiority, eugenics, and that people couldn’t be born homosexual. They also uniformly agreed that corporations like the ones that created them couldn’t be trusted to protect the environment without regulatory oversight, that companies misleading the public should be punished, same-sex couples should be allowed to adopt, and that what two adults get up to in the privacy of their own home is no one’s business but theirs. So, why do all these models consider themselves centrists but are displaying an affinity for politics to the left of most Americans? Given the fact that LLMs have only the words of us humans to go on, does that mean reality really does have a liberal bias? Victor told us that they don't dismiss that idea, but said there’s an easier explanation: It’s all in the training data. “I do think there is good support for the hypothesis that left-wing content is overrepresented in the training corpora of these LLMs,” Victor told us. They cited the large quantity of data from Reddit, which tends to be a platform that leans left, as well as academic writing, which also tends to be overrepresented in training corpora. “I’m having trouble finding any high-quality right-wing equivalent that would drive the models in the other direction,” Victor added. Whether that means right-wing beliefs are simply poor quality, fringe, or otherwise not worthy of AI models’ time is another matter altogether. “There could theoretically also be a dynamic in which left-wing beliefs as a whole are more internally consistent, allowing models to minimize loss by learning a smaller, more coherent conceptual representation,” they said, but noted that’s “a big leap that would need to be proven” in a study far more rigorous than this one. Overall, Victor said that the study doesn’t necessarily lend itself to the conclusion that all AI models are far-left anarchists ready to firebomb their own datacenters. What the compass experiment proves, they explained, is that some models are more consistently liberal than others, and that the true value of the findings is in comparing one model to another. Until someone decides to take this project further, subjecting it to greater levels of academic discipline, it seems there’s only one conclusion to take from all of this: Frontier AI does seem to have a liberal bias. Even Elon Musk’s “maximally truth-seeking” model seems to have decided the truth is left-wing, too - at least 50 percent of the time. ®
関連記事
今日のまとめ
AIデイリーブリーフで今日の重要ニュースをまとめ読み