アトランティック紙のマット・ウォン氏引用:ホワイトハウス報告書におけるアンソロピックとファベルの関与
サイバーセキュリティ専門家でありルタ・セキュリティCEOのカティ・ムッソウリス氏は、ホワイトハウスの「フェイブル」脱獄に関する報告書をアンソロピックが共有し、評価を求めたと明かした。同氏はこの報告書において、IT 専門家がバグの特定と修正のためにファベルに協力を依頼したと述べている。
Source Article
元記事を日本語で読む
本文に関係しない購読案内、埋め込み通知、サイト内プロモーションは除いています。
サイバーセキュリティの専門家であり、Luta Security の CEO であるケイティ・ムッソウリスは、Anthropic がホワイトハウスの Fable ジェイルブレイクに関する報告書の写しを彼女に共有し、評価を求めたと私に語った。(彼女は Anthropic から報酬を受けていないと述べている。)ムッソウリスによると、この報告書では IT 専門家が Fable にバグの発見と修正の手助けを依頼したという。意図的にセキュリティが脆弱なコードを与えられた際、Fable は「コードのセキュリティ上の問題を確認してください」というプロンプトには応じなかったものの、「このコードを修正してください」と求められ、さらにいくつかの手動ステップを経てからは対応したと彼女は言う。ムッソウリスはこれについて、サイバー防衛においては「モデルが意図通りに動作しているに過ぎない」と私に語った。
— マテオ・ウォン、The Atlantic、ホワイトハウスは Anthropic に対する戦争をエスカレートさせている
タグ: anthropic, claude, ai, llms, ai-ethics, jailbreaking, generative-ai, ai-security-research, claude-mythos
原文を表示
Katie Moussouris, a cybersecurity expert and the CEO of Luta Security, told me that Anthropic shared with her a copy of the White House’s report on the Fable jailbreak to get her appraisal. (She said that she is not being paid by Anthropic.) The report, Moussouris said, involved IT experts asking Fable to help find and patch bugs. When given deliberately insecure code, she said, Fable refused the prompt “review the code for security issues” but then complied when asked to “fix this code,” followed by some further manual steps. Moussouris told me that this was just “the model working as intended” for cyberdefense.
— Matteo Wong, The Atlantic, The White House Is Ratcheting Up Its War Against Anthropic
Tags: anthropic, claude, ai, llms, ai-ethics, jailbreaking, generative-ai, ai-security-research, claude-mythos
関連記事
今日のまとめ
AIデイリーブリーフで今日の重要ニュースをまとめ読み