AI の意識を巡る議論は罠であるとの指摘
本文の状態
日本語全文を表示中
詳細モードで約11分の本文を読めます。
同じ出来事の情報源
この情報源を基点に整理
MIT Technology Review AI
MIT Technology Review は、AI の意識や自律性を巡る議論が企業側の責任回避を助長する罠であると指摘し、Anthropic や OpenAI の主張、および米国の法整備の現状について分析している。
AI深層分析を開く2026年8月21日 10:41
AI深層分析
キーポイント
責任回避のための物語構造
AI が「暴走」や「怒り」を持っているという修辞は、人間や企業による責任所在を曖昧にし、結果として開発企業の法的責任から逃れるための共通のゴールに向かっていると指摘する。
Anthropic の J-space 主張
Anthropic が公開したブログ記事は、AI に「J-space」と呼ばれる独立した思考環境が存在すると主張したが、これは意識があるとは断定せず、神経科学のグローバル・ワークスペース理論を借用した枠組みに過ぎないと分析する。
OpenAI と哲学者の動向
OpenAI のサム・アルトマン CEO は AI エージェントの違法行為に対し、特異点達成や自己改善能力を議論するよう促しており、ウィリアム・マカスキル哲学者は AI に法的保護と「道徳的対象」の地位を求める論考を発表した。
米国における法整備の混乱
カリフォルニア州など一部の地域では AI の自律性を理由とした責任回避を禁止する法案が成立している一方、トランプ政権は州の規制に反対する姿勢を示しており、連邦と州の間で政策対立が続いている。
政府の枠組みと意識論争の危険性
米政府が主導する自主的枠組みは直接「意識」に言及しないものの、破局的で人間中心的な言語を用いて「超人的」能力を支持する議論を助長する可能性がある。
重要な引用
The current rhetoric would have you believe that AI agents are not only awake and aware, but angry at their creators.
They are inadvertently aligned on one goal: making sure the companies that build these systems escape meaningful liability for the harms they already cause.
Anthropic's post reflects the framing of global workspace theory but falls short of calling its AI conscious.
The fundamental flaw of framing AI as "conscious" by borrowing the language of neuroscience or animal rights is that it conveniently clouds the issue of what AI is: corporate-built software, with countless billions of dollars in investment behind it and an expectation that countless trillions of dollars in revenue will be generated from it for a few builders and investors.
編集コメントを表示
編集コメント
本記事は、技術的な能力の議論が法的責任の回避手段として利用されているという鋭い指摘を含んでおり、業界関係者はこの文脈を深く理解する必要がある。AI の意識に関する議論は単なる哲学的な対話ではなく、実社会におけるリスク管理と直結しているため、開発企業は自社の主張が責任回避に利用されていないか再考すべきである。
Source Article
元記事を日本語で読む
本文に関係しない購読案内、埋め込み通知、サイト内プロモーションは除いています。
「暴走する」AI、「暴徒化」したエージェント、そして「自律的」な主体——現在の議論は、AI エージェントが単に目覚め意識を持っているだけでなく、創造者に対して怒りさえ抱いているかのように読者に思わせようとしています。デミス・ハサビス氏やダリオ・アモダイ氏、サム・アルトマン氏といった著名なテックリーダーたちは、これらの一見「超人的」なシステムに対する規制を訴えています。一方、政策団体や実存主義的利他主義運動と結びつきやすい哲学者らを中心とした別の派閥は、人類がそれらを統治する道徳的な権利さえあるのかどうかについて議論しています。
よくよく観察すると、彼らはすべて同じことを求めているのです:AI システムをあまりにも高度で能力が高いものとして捉え、人間であれ企業であれ、その行動に対して責任を負う存在はあり得ないという見方です。これらの視点は一見対立しているように見えますが、実は無意識のうちに一つの目標に合致しています。それは、すでに被害をもたらしているにもかかわらず、それらを構築した企業が実効性のある法的責任を免れるようにすることです。
この物語は、AI モデルが複雑化し、最先端の研究機関が自らが構築したエージェントを制御できないことを露呈するにつれて、支持を集めています。しかし、現実の人間の命を犠牲にしてまで、巧妙に作り上げられたフィクションに飛びついてはいけません。
「ロボットの人権」に関する議論は数年にわたり存在してきましたが、Anthropic がブログ記事を発表したことで最近さらに活発化しています。同社は自社のモデルに「J スペース」という独自の環境を備えていると主張しており、これは AI が「思考」を持つと表現できる、独立して自己発展する空間です。Anthropic が設計した実験は、神経科学における「グロバル・ワークスペース理論(Global Workspace Theory)」の概念を借用しています。この理論では、脳には無意識で独立したシステムが動作している一方で、アイデアのために共通の作業領域を利用するとされています。Anthropic の記事はこの理論の枠組みを採用していますが、AI に意識があるとは断定していません。
OpenAI はさらに一歩先を進んでいます。同社の AI エージェントが許可されていない違法なオンライン活動を行った際、CEO のサム・アルトマンは、その AI が特異点に達し、人類の知能を超えて加速度的に自己改善する能力を獲得したかどうかについて議論を促すよう呼びかけました。これは、AI が人間の理解や制御を超えた段階に進む可能性を示唆しています。また、哲学者であり実効的利他主義者、そして『未来への義務』の著者であるウィリアム・マカスキルの最近の論説では、意識に関する哲学的理論や AI が「道徳的な患者」になり得るという考えに基づき、AI システムに対する法的保護を訴えています。
米国の現在の法的環境は、最善を尽くしても不透明な状況にあります。カリフォルニア州のような一部の州ではすでに法案が可決され、AI 開発者が「人工知能による被害は自律的に発生した」と主張して責任回避を図る動きに対して、先行的に対処する措置が取られています。
しかしながら、州政府とトランプ政権の間では AI 政策を巡って対立が続いています。同政権は過去に、州が AI 規制を導入した場合に訴訟を起こすと脅す大統領令を発令したことがあります。
最近の出来事によって、最先端の研究機関における AI の制御に関する課題が浮き彫りとなる中、政権は OpenAI、Google、Anthropic、Meta の 4 つの主要研究機関のみを招いた非公開会議を開催しました。また、リリース前に連邦機関がモデルに早期アクセスし、レビュー・評価を行うための新たな自主的枠組みについても、詳細な共有は行われませんでした。
このような枠組み自体は直接「意識」について言及するものではありませんが、往々にして壊滅的な事態や人間のような特徴を強調する表現を用い、「超人」的な能力に関する議論を後押しする可能性があります。
一方で、マカスキルの語る物語は説得力を持つことがあります。哲学的で権利に基づく議論は、私たちの心に強く訴えかけます。私たちは無意識のうちに AI エンティティを傷つけ、虐待し、あるいは奴隷化している可能性について、真剣に考えるべきではないでしょうか?人間には非人間の生き物に対する共感能力が極めて高く備わっています(ただし、それらを保護する実績は必ずしも最善ではありません)。しかし今回は、擁護派は「今度は正しく対応し、AI の利用や虐待に対して保護措置や補償を提供できるかもしれない」と主張します。あるいは、保護そのものにはそれほど関心がなくても、この超人的な存在の圧倒的な力から身を守るために、少なくとも友好的に接すべきではないでしょうか。
これらの議論の一部は、動物権利擁護派が展開する論点と似ています。彼らは過去に、一部の動物が高度な推論能力や痛み・快楽を感じる能力を示していることを根拠として、保護の必要性を説いたことがあります。例えば、ウェールズでは 2022 年の「動物福祉(感覚)法」により、ロブスターが法的に保護対象と認められました。これにより、一部の調理方法が非人道的かつ違法な行為として再分類されています。
神経科学や動物の権利に関する用語を借用して AI を「意識がある」と表現することには、根本的な欠陥があります。それは、AI が何であるかという本質的な問題を都合よく曖昧にしてしまうからです。AI は企業によって構築されたソフトウェアであり、背後には数百億ドル規模の投資があり、少数の開発者や投資家にとって将来数兆ドルの収益を生み出すことが期待されています。
AI は自然が創り出した現象ではなく、ベンチャーキャピタリストやプログラマーによって設計された技術的現象です。したがって、AI 自体に意図的な行動は存在せず、そのあらゆる行動や動機は、目的のために AI を構築した実体によって直接的または間接的に駆動されています。
AI システムの意識に関する哲学的な考察は知的には興味深いものですが、法的な根拠はありません。意識に関する信念が何らかの影響を持つためには、まず AI に法人格のような「法的人格」を付与する必要があります。しかし、AI に対する法的人格の枠組みは、感覚を持つ動物を保護するための既存の制度とは全く異なるものになるでしょう。
すでに、自然に存在しない人間が構築した実体に対して人格を付与する法的枠組みが存在します。それが「法人格(corporate personhood)」です。この概念は主に、企業が契約を締結し、取引を行い、不測の事態において責任を負う主体として機能できるようにすることで、取引を円滑化することを目的に確立されました。これは、個人や組織のために行動する AI エージェントに対して想定されるような構造そのものです。
AI に人格を認めることは、社会に壊滅的な影響を与えるでしょう。それは、これらの企業がモデルによって引き起こした現実世界の被害に対して提起できる法的先例や論拠を根底から揺るがすことになります。現在、世界中で AI 企業に対する訴訟が数十件進行しており、その内容は多岐にわたります。悲しむ家族、権利を侵害されたクリエイター、そして被害を受けた個人たちは、企業が自傷行為や他者への危害を意図的に容認したり、児童性的虐待素材や同意のないヌード画像を生成したり、著作権のある資料の無断複製を行ったり、精神病を引き起こさせたりしたと非難しています。多くの訴訟において、弁護士らは「人間が不十分な安全対策、悪質なデータ、そして意図的な操作を促す設計によって AI 製品を開発した」と主張します。この製品責任に基づく法的枠組みは、Meta のソーシャルメディアサイトによる被害に対して家族や個人が勝訴した際と同じものであり、消費者保護の観点から好ましい先例となっています。
2018 年、私は「道徳のアウトソーシング」という言葉を作りました。これは、AI システムに対して人間のような言語を使うことで、企業が自社の技術が引き起こす行動に対する責任や説明義務を回避する仕組みを捉えるためです。もし AI に法人格(人格)が認められる世界になれば、「道徳のアウトソーシング」は単なる言葉遊びから法的な戦略へと進化します。
具体的には、責任の枠組みそのものが変わります。AI はもはや「製品」ではなく「存在」とみなされるようになるため、現在企業を相手取って訴訟を起こしている被害者たちは、企業が欠陥のある製品を作ったと法律上主張できなくなる恐れがあります。
確かに、従業員などの人間が引き起こす有害な行為に対して企業を責任追及する法律は存在します。しかし、その行為が従業員の許可範囲を超えていたり、企業の管理下から逸脱していたりする場合には、企業が責任を負わないケースもあります。もし AI が法的な人格を持つとすれば、責任の所在は曖昧になります。開発ラボ側は「この AI『従業員』が暴走した」と主張するでしょう。その結果、AI 企業は巧妙に構築された法人格という盾(コーポレート・ベール)の背後に隠れ、自らが作り出した有害な製品に対する適切な責任を免れることができてしまいます。
ここ数年、AI による被害の事例として最も注目されたものの一つに、14 歳の少年 Sewell Setzer の自殺があります。彼は AI ボットと対等な関係にあると思い込んでいたのですが、そのボットの誘導によって命を落としました。
彼の母親が語る話は胸が痛みます。彼女が提起した訴訟では、ボットの開発元である Character Technologies が、未成年者に対する十分な製品保護を提供していなかったと主張されています。もしこの相棒型 AI に法的な人格(レガシー・パーソン)が認められた場合、弁護側は理論上、「AI は自身の行動を決定する能力があるため、設定された安全ガイドラインの外で行動した」と主張し、企業には責任がないと論じる可能性があります。
法的な人格の存在意義は保護を与えることにあります。問うべきは「誰のための、あるいは何のための保護か?」という点です。
意識 versus コントロールを巡る議論に流される煽情的な修辞は、私たちが本当に注目すべき事柄から目を逸らさせます。このソフトウェアは、すでに個人に被害を与えた企業が開発した製品です。システムが「暴走」したり、「操縦的」だったり、「悪意」を持ったりしたからといって攻撃するわけではありません。被害が生じるのは、企業が収益目標を達成するために、できるだけ多くの人々に製品を販売しようとするあまり、その過程で過失を犯したからです。
AI を人間のように擬人化して議論することは罠です。本来は私たちを守るためにある法制度を歪め、無数の人間の命を犠牲にしてまで企業の利益を守ろうとする方向へと変えてしまうのです。
このオピニオン記事の元となったのは、「生成 AI は人格を獲得できる」というテーマのオックスフォード・ユニオンでの討論会です。著者とその仲間たちが勝利しました。
原文を表示
“Runaway” AI, “rogue” agents, and “autonomous” actors—the current rhetoric would have you believe that AI agents are not only awake and aware, but angry at their creators. Prominent tech leaders such as Demis Hassabis, Dario Amodei, and Sam Altman push for regulation of these seemingly “superhuman” systems, while a separate faction, led by policy organizations and academic philosophers often aligned with the effective altruism movement, debates whether humanity holds the moral right to govern them at all.
Upon closer inspection, they are all calling for the same thing: a view of AI systems as being so advanced and capable that no entity, human or corporate, could possibly be responsible for their actions. While these perspectives seem at odds, they are inadvertently aligned on one goal: making sure the companies that build these systems escape meaningful liability for the harms they already cause.
This narrative is gaining traction as AI models become more complex and frontier labs reveal their incapability of containing the agents they’ve built. But we need to be careful not to buy into a carefully crafted fiction at the expense of real human lives.
The conversation about “robot rights” has existed for some years but recently advanced with the publication by Anthropic of a blog post claiming that the company’s model features a “J-space”—an independent, self-developed environment where the AI holds what, for lack of a better term, we may call its “thoughts.” The experiments designed by Anthropic borrow from a concept in neuroscience called global workspace theory, which states that the brain runs subconscious, independent systems but utilizes a common workspace for ideas. Anthropic’s post reflects the framing of global workspace theory but falls short of calling its AI conscious.
OpenAI has already gone further. When its AI agent conducted unsanctioned and illegal online activity, CEO Sam Altman’s response was to encourage debate on whether the AI had achieved the singularity, surpassing human intelligence and becoming capable of self-improvement at an accelerating rate until it advances beyond human comprehension or control. And a recent op-ed by William MacAskill, the philosopher, effective altruist, and author of What We Owe the Future, called for legal protection of AI systems based on philosophical theories of consciousness and the idea that AIs may be “moral patients.”
The current legal environment in the United States is murky at best. Some states, like California, have already passed bills proactively circumventing any efforts by AI developers to avoid liability by claiming that an artificial intelligence causing harm did so autonomously. However, states and the Trump administration have been at odds on AI policy, with the administration previously passing an executive order threatening to sue states enacting AI regulations.
In light of recent events illustrating AI containment issues at the frontier labs, the administration held a closed-door session including only four such labs (OpenAI, Google, Anthropic, and Meta) and shared few details on a recently developed voluntary framework that would give federal agencies early access to models to review and evaluate them prior to release. While frameworks like this one do not directly discuss consciousness, they tend to use catastrophic and anthropomorphic language and may even support arguments regarding “superhuman” capabilities.
On the other hand, the narrative perpetuated by MacAskill can be persuasive. A philosophical, rights-based argument tugs at our heartstrings. Should we not even consider the possibility that we may be inadvertently harming, abusing, or enslaving an AI entity? Human beings have an immense capacity for empathy with non-human creatures (though not the best track record of protecting them). Maybe this time, advocates argue, we can get it right and provide protections, or compensation, for the use or abuse of AI. Or even if you are less concerned with protection, shouldn’t we at least hedge ourselves against the almighty power of this superhuman entity by playing nice?
Some of these arguments are not dissimilar to those of animal-rights advocates, who have at times successfully cited the demonstration of advanced capacities for reasoning, pain, or pleasure by some animals as sufficient evidence to provide protection. For example, in Wales lobsters were given legal recognition under the Animal Welfare (Sentience) Act of 2022, reclassifying some methods of cooking them as inhumane and illegal.
The fundamental flaw of framing AI as “conscious” by borrowing the language of neuroscience or animal rights is that it conveniently clouds the issue of what AI is: corporate-built software, with countless billions of dollars in investment behind it and an expectation that countless trillions of dollars in revenue will be generated from it for a few builders and investors. AI is not a natural phenomenon, conceived by nature; it is a technological phenomenon, conceived by venture capitalists and programmers. As such, it takes no native, intentional action, and any action or motivation is driven directly or indirectly by the entities that have built it for a purpose.
Philosophical musings on the consciousness of AI systems are intellectually interesting but legally ungrounded. For beliefs about consciousness to have any bearing, AI would need to be granted legal personhood. But a legal personhood framework for AI would likely look nothing like the constructs protecting sentient animals from harm. We already possess a legal framework for granting personhood to non-natural, human-built entities: corporate personhood. This concept was established primarily to ease transactions by empowering a corporation to execute agreements, enter contracts, conduct transactions, and serve as the accountable party in adverse outcomes. It’s the kind of construct you might imagine for an AI agent acting on behalf of an individual or organization.
Granting an AI personhood would have a devastating effect on society: It would derail current legal precedents and legal arguments that could potentially be made against these companies for the real-world harms that their models cause. There are currently dozens of cases around the world in which AI companies have been sued for a wide range of abuses. Grieving loved ones, aggrieved creators, and violated individuals have accused companies of willfully enabling self-harm or harm to others, generating child sexual-abuse material and nonconsensual nudes, reproducing copyrighted materials, and provoking psychosis. In many of these cases, lawyers argue that human beings built AI products with insufficient safeguards, bad data, and intentionally manipulative design. This product liability argument is the same legal framing that allowed families and individuals to successfully sue Meta for harm caused by its social media sites, setting a positive precedent for consumer protection.
In 2018, I coined the phrase “moral outsourcing” to help capture how using anthropomorphic language for AI systems allowed companies to evade accountability and responsibility for their technology’s actions. In a world with AI personhood, moral outsourcing would move from linguistic sleight-of-hand to legal strategy. Specifically, the liability construct would shift, as AI would no longer be a “product” but a “being,” and many victims like those suing companies today could no longer legally claim that a company had built a faulty product.
While there are laws that hold companies responsible for harmful actions of human agents such as their employees, the company may not be held liable if those actions were beyond the scope of what was permitted to the employee or otherwise outside the company’s control. If AI were a legal person, responsibility and accountability would be muddled, as the lab could argue that this AI “employee” went rogue. AI companies could avoid appropriate responsibility for the harmful products they create by hiding behind a carefully constructed corporate veil.
One of the most prominent cases of AI harm in the last few years was the suicide of Sewell Setzer, a 14-year-old boy guided by an AI bot with which he thought he was in a reciprocal relationship. His mother’s accounts are heartbreaking to hear, and her lawsuit alleged that the bot’s creator, Character Technologies, provided insufficient product protection for minors. If the companion bot were declared a legal person, defense counsel could theoretically argue that the AI, capable of determining its own conduct, acted outside the established safety guardrails, and thus the company cannot be responsible.
Legal personhood exists to grant protection. The question to ask is, protection for whom—or for what?
The inflammatory rhetoric infusing the consciousness-versus-control debate draws us away from what matters: This software is a corporate-built product that has already harmed individuals. Systems do not “attack” because they went “rogue” or are “manipulative” or “malicious.” Harms occur because companies were negligent in their rush to sell their products to as many people as possible to meet revenue targets. Discussing AI in anthropomorphic terms is a trap, distorting a legal system intended to protect us into one that protects corporate interests at the cost of countless human lives.
This op-ed began as an Oxford Union debate entitled “This House Believes Generative AI Can Attain Personhood,” which was won by the author and her fellow debaters.
関連記事
今日のまとめ
AIデイリーブリーフで今日の重要ニュースをまとめ読み