この AI スタートアップは、脚本がヒット映画になるかを予測できると主張している
映画脚本のヒットを予測するAIスタートアップ「Quilty」は、実際のテストで失敗作を成功作より高く評価するなど精度に疑問視されており、業界への実用性は現時点では不透明である。
キーポイント
予測精度の実証的欠如
Quilty のツールは、後に興行不振となった映画の脚本を成功作よりも高く評価するなど、実際のデータに基づいたテストで信頼性に疑問を持たれている。
既存技術の組み合わせに過ぎない
業界関係者からは、Quilty が独自の高度な分析能力を持つのではなく、既存のAIシステムを単に組み合わせたものに過ぎないと批判されている。
人間中心のアプローチと目的
創業者は完全自動化ではなく「人間の創造性を支援する」ツールとして位置づけ、脚本家やプロデューサーが意思決定を行うための情報提供を主眼としている。
重要な引用
Even with all the available data in the world, it predicted the script for Christy... would outperform the script for Sinners
Quilty is little more than a jumbled mishmash of preexisting AI systems
We agree with a lot of the negative sentiment towards AI, but what we're trying to do is enable human creativity
影響分析・編集コメントを表示
影響分析
この記事は、生成AIが映画制作のようなクリエイティブ分野で「予測ツール」として実装される際の限界とリスクを浮き彫りにしています。特に、過去のデータに基づく学習モデルが、芸術的な成功や市場の複雑な変動を正確に捉えられない可能性を示唆しており、業界関係者がAIツールの導入を慎重に行う必要があることを警告しています。
編集コメント
AIがクリエイティブな判断を代替できるという楽観論に対し、実証データによる懐疑的な視点が示された重要なケーススタディです。技術の進歩よりも、その適用範囲と限界を理解することが業界にとっては重要です。
今年初めに業界紙に登場した際、AI スタートアップの Quilty は、脚本を読むだけで映画の成功を正確に予測できるツールを提供すると約束しました。しかし、実際に Quilty の製品を試す機会 を得た人々は懐疑的になりました。世界中のあらゆるデータが存在するにもかかわらず、このツールは後に興行の大失敗となることが判明した Christy の脚本 が、アカデミー賞を受賞した大ヒット作となった Sinners の脚本 よりも優れた結果を出すと予測してしまったのです。
多くの AI 経営陣が以前に主張してきたように、Quilty の創設者たちは、このツールが新興のクリエイターたちに支援ツールへのアクセスを提供することで業界を「民主化」できると信じています。おそらく Quilty からの高いスコアはプロデューサーとの接点となり、低いスコアはさらなる修正が必要であるというシグナルとなるでしょう。しかし現在、Quilty は既存の AI システムがごちゃ混ぜになったものに過ぎず、同社はまだ自社の技術に将来の大ヒット作(ましてや既に証明された作品)を特定する審美眼や分析能力があることを示すことができていません。
映画プロデューサーのサイモン・ホースマン(Simon Horsman)とダニエル・ウッド(Daniel Wood)によって設立された Quilty は、AI を活用して脚本を分析し、プロジェクトの成功確率に関する詳細なレポートを生成します。未製作の脚本を入力すると、Quilty の技術は 0 から 100 のスコアを付与し、そのスコアには物語の質、商業的な実現可能性、観客との共鳴度、そして制作にかかる予想コストが含まれます。このプラットフォームは、映画や作品が承認される過程で、ユーザーに未来の一瞥を提供できると訴求しています。ホースマンとウッドは、Quilty が従来の製作スタジオのビジネスプロセスにおいて不可欠な要素となる準備ができていると考えています。
最近、私はホースマンとウッドと面談しましたが、彼らは前工程を完全に自動化するのではなく、「人間をループ内(in the loop)に保つ」ことを強く望んでいました。会社設立当初、ホースマンとウッドは多くのクリエイターからフィードバックを求めましたが、その多くが生成 AI が雇用を悪影響を与え、人間の労働者のスキルを低下させる可能性について懸念を表明していました。
「AI に対する否定的な感情の多くには同意しますが、私たちが目指しているのは人間創造性の発揮です」とホースマンは語りました。「Quilty は開発に焦点を当てたものであり、脚本家、プロデューサー、バイヤー、資金提供者、スタジオ役員といったユーザーに対して、情報に基づいた承認判断を下せるよう、可能な限り多くの情報を提供することを目指しています。」
Quilty は、スクリプトに対してフィードバックを与える単一の専用 AI モデルへのアクセスを提供するのではなく、さまざまな分析をプロセスに持ち込むために、広く利用可能な複数の AI ツールを組み合わせています。ユーザーがやるべきことは、テキストスクリプトをプラットフォームにアップロードすることだけで、数分後には、推定予算や重要なストーリーの要点の概要、キャラクター分析などを詳細に記載したレポートが出力されます。このサービスの料金は 1 回あたりの分析で 50 ドルですが、複数の分析を割引価格で購入することも可能です。
このような断片的な分析ワークフローのアイデアは、Quilty の CTO も務めるウッド氏が数年前に不動産問題で訴えられた際に初めて浮かびました。弁護士費用を支出する代わりに、ウッド氏は ChatGPT を起動し、すぐに「私は弁護士ではありません; 私以外の誰かを探してください」と言われました。
「その後、Gemini に移りましたが、コンテキストウィンドウが大きかったためしばらくはうまく機能しました」とウッド氏は振り返ります。「しかし、X で 愚かなイーロン・マスク が AI モデルとして最高の弁護士スコアを獲得した Grok について話しているのを見て、『確かめてみよう』と思いました。」(ウッド氏はその法的紛争がどう決着したかについては詳細を明かさなかった。)
「Claude Mythos が登場すれば、突然、私のソフトウェア全体が向上する」
この経験を通じて、ウッドは同様の消費者向け AI モデルが異なるタスクにおいてどのように優れているかをより深く理解しました。また、ウッドの個人的な AI 利用経験が、脚本の可能性を定量化する Quilty のアプローチにも反映されています。「Gemini は構造とパターンに対して素晴らしいツールです」とクィルティは述べ、このツールを使用して映画や番組の制作要素を包括的なリストに要約した文書(ブレイクダウン)の生成を支援しています。財務モデリングについては、米国に設置されたサーバー上でホストされている DeepSeek のインスタンスを信頼して利用しています。一方、物語やキャラクター分析には、Claude と ChatGPT を組み合わせて使用しています。
ウッドは、ハルシネーション(幻覚)に満ちた出力ではなく質の高い結果を生成するために、同社が文脈プロンプティング(追加の文脈データを提供して処理を行うプロセス)に依存していると私に語りました。クィルティは、映画レポートやスコアを作成するために使用するモデルを個人的にトレーニングしたことはありません。しかしウッドは、これは弱点ではなく強みであると強調しました。その理由は、新しいおよび改良されたモデルが一般公開されるたびに、Quilty がそれらをワークフローに取り込みやすくなるからです。
「Claude Mythos が登場し、それがより優れた大規模言語モデル(LLM)であることが確認できれば、突然私のソフトウェア全体が向上します」とウッド氏は述べました。これはサイバーセキュリティ目的のために限られた組織のみが利用可能な 強力な新モデル を指しています。「もし中国製のモデルが突然、これらの米国最先端モデルすべてよりも優位になれば、なぜ私はそれらを使わないのでしょうか?」
Quilty の技術スタックのモジュール性(modularity)は、全体的なアップデートにおいてはより俊敏性を発揮する可能性がありますが、同時に、このプラットフォームが脚本をどのように受け取り、まだ存在しない映画に対する観客の反応など、目に見えない要素を測定するとされる一連の指標をどう導き出すのかを完全に理解することが、やや難しくなる側面もあります。予測はハリウッド誕生以来、映画開発における重要な要素でしたが、その業務は従来、観客への微妙な理解を持つ人間労働者によって行われてきました。
AI 企業はいまだ、人間の思考プロセスを真に再現したり、芸術に対する我々の不確かな意見形成の仕方を模倣したりするモデルを開発できていない。しかし、Quilty の創設者たちは、同社の「センチメントエンジン」が脚本の評価において次善の策であると信じている。その理由は、VADER(Valence Aware Dictionary and sEntiment Reasoner)のようなツールを組み込んでいる点にあるからだ。VADER はオープンソースソフトウェアであり、テキストがポジティブかネガティブかにどの程度傾いているかを測定するものである。
Quilty が映画の受け入れ方に影響を与えるあらゆる要因をすべて予測することは不可能である。
Horsman 氏と Wood 氏はまた、Quilty がプロジェクトが「文化的な瞬間」にどのように対応しているかを正確に特定でき、信頼性の高い興行収入予測を提供できると確信している。彼らは、『ネバー・エンディング・ストーリー』のような人気のある古い映画を例として挙げた。この映画は、性的暴行をコメディ調で描こうとする点(現代の視聴者にとっては品位に欠けると見なされる行為)が原因で、Quilty のスコアが低くなるだろうと指摘した。
ホースマンとウッドに、なぜクィルティが『クリスティ』(最終的に興行収入は約 200 万ドル)を『サイナーズ』(興行収入 3 億 7000 万ドル)よりも高いスコアで評価したのかを尋ねた際、彼らはプラットフォームの判断は「シドニー・スウィーニーが非常に、非常に人気がある」という事実に帰着すると主張しました。彼らによると、紙面上では、スウィーニーのスターパワーと、ボクシングに関する伝記ドラマが『サイナーズ』のようなファンタジー/アクション映画よりも制作コストが安いという事実が組み合わさり、『クリスティ』の方が安全な賭けになると判断されたそうです。しかし、この状況はクィルティのロジックがそれほど信頼できるものではないことを浮き彫りにしています。ホースマンとウッドは、クィルティが映画の財務パフォーマンスや観客の受け入れ方に影響を与える可能性のある要因を予測できない場合があることを認めました。
例えば、クイリーは、エライジャ・バイナムの『マガジン・ドリームス』(これはホースマンがプロデュースした作品)が、2023 年に俳優ジョナサン・メイジャーズの目立った失脚によって頓挫するとは、決して予想できなかったでしょう。同様に、『マインクラフト』映画の脚本が、チキン・ジョッキー現象が映画の大成功の一部となることを示唆していたわけではありません。ホースマンとウッドは私に、最終的にはクイリーもこうした出来事を予測できるようになることを望んでいると話しましたが、それがどのように実現するかなど想像するのは難しいものです。
多くの注目を集めているにもかかわらず、クイリーが提供しているのは、未制作の芸術作品に関する未来を予測するように要請された、さまざまな大規模言語モデルへの回りくどいアクセスに過ぎません。もしこれらの AI ツールのいずれかが、クイリーが主張するような形で機能するのであれば、それは本当に驚異的なことでしょう。しかし、それらのほとんどは単なる高度なパターン認識/模倣機械であり、人間が何を面白がるかを理解するに至るには、まだ遥か遠くにあります。
この物語のトピックや著者をフォローして、パーソナライズされたホームページフィードで類似の記事をもっと見たり、メール更新を受け取ったりしましょう。
- チャールズ・パリヤム=ムーア
-
-
-
原文を表示
When Quilty hit the industry trades earlier this year, the AI startup promised that its tool could accurately predict a film’s success just by reading the script. When people actually got a chance to experiment with Quilty’s product, though, they were left skeptical. Even with all the available data in the world, it predicted the script for Christy, which would go on to bea box office flop, would outperform the script for Sinners, which became an Oscar-winning blockbuster.
As many AI execs have pitched before, Quilty’s founders believe that can help “democratize” their industry by giving up-and-coming creatives access to assistive tools — a great Quilty score, perhaps, could be an in with a producer, and a low score might be a sign more revisions are needed. But right now, Quilty is little more than a jumbled mishmash of preexisting AI systems, and the company has yet to prove out that its technology has the taste or analytical abilities to identify a future hit (let alone a proven one).
Founded by film producers Simon Horsman and Daniel Wood, Quilty uses AI to analyze scripts and generate detailed reports about a project’s chances for success. After being fed an unproduced script, Quilty’s tech gives it a score ranging from 0 to 100 that reflects the quality of the would-be project’s narrative, its commercial viability, whether it will resonate with audiences, and how much the production would likely cost. The platform is selling the idea that it can give users a glimpse into the future as they try to get their films / movies greenlit. Horsman and Wood believe that Quilty is poised to become an integral part of how traditional production studios do business.
When I recently sat down with Horsman and Wood, they were adamant about wanting to “keep humans in the loop” rather than fully automating the pre-production process. While first establishing their company, Horsman and Wood solicited feedback from a number of other creatives who often voiced concerns about gen AI’s potential to negatively impact jobs and leave human workers deskilled.
“We agree with a lot of the negative sentiment towards AI, but what we’re trying to do is enable human creativity,” Horsman told me. “Quilty is really about development and giving the users — be they a writer, producer, buyer, financier, or studio execs — as much information as possible to make an informed green light decision.”
Instead of offering users access to a single, bespoke AI model that gives feedback on scripts, Quilty combines a number of widely available AI tools to bring different kinds of analyses to the process. All users have to do is upload their text scripts to the platform, and a few minutes later, it spits out a report that details things like an estimated budget, outlines of important story beats, and character analyses. The service costs $50 per individual analysis, but you can also purchase multiple analyses at a discounted rate.
The idea for this kind of piecemeal analytical workflow first came to Wood — who also serves as Quilty’s CTO — a few years back when he was being sued over a real estate matter. Rather than spending money on an attorney, Wood fired up ChatGPT, which promptly told him “I’m not a lawyer; go find someone else to help you.”
“Then I went to Gemini, which worked a lot better for a while, because I had a larger context window,” Wood recounted. “But then I was on X, and saw stupid Elon Musk talking about Grok getting the best lawyer score ever for an AI model, and I was like, ‘Let me check that out.’” (Wood did not detail how that legal dispute worked out.)
“When Claude Mythos comes out, all of a sudden, my whole software gets better”
The experience left Wood with a better understanding of how similar consumer-grade AI models can excel at different tasks. And Wood’s personal use of AI has informed Quilty’s approach to quantifying a script’s potential success. Because “Gemini is fantastic for structure and patterns,” Quilty uses that tool to help generate breakdowns — documents that distill all of a film or show’s production elements into comprehensive lists. For financial modeling, the company puts its faith in an instance of DeepSeek that’s hosted on servers located in the US. And for narrative / character analysis, Quilty uses a combination of Claude and ChatGPT.
Wood told me that the company relies on context prompting — a process in which you provide additional contextual data — in order to generate quality outputs that aren’t filled with hallucinations. Quilty doesn’t personally train any of the models it uses to create film reports / scores. But Wood insisted that it was a strength rather than a weakness because it makes it easier for Quilty to incorporate new and improved models into its workflow as they become available to the public.
“When Claude Mythos comes out and I can see that it’s a better LLM, all of a sudden, my whole software gets better,” Wood said, referring to the powerful new model that’s only available to a small group of organizations for cybersecurity purposes. “If some Chinese models suddenly become better than all these US frontier models, why wouldn’t I just use those instead?”
Though the modularity of Quilty’s tech stack might make it more agile in terms of overall updates, it also makes it somewhat harder to fully understand how the platform takes a script and comes up with a bevy of metrics that purportedly measure intangible things like how an audience *might *react to a movie that doesn’t actually exist yet. Prediction has been a key part of film development since the birth of Hollywood, but that labor has traditionally been performed by human workers who have a nuanced understanding of audiences.
No AI firm has been able to develop a model that can truly replicate human thought processes or the imprecise way we form opinions about art. But Quilty’s founders think that their “sentiment engine” is the next-best thing when it comes to assessing scripts because of the way it incorporates tools like VADER (Valence Aware Dictionary and sEntiment Reasoner) — open-source software that measures the degree to which text comes across as positive versus negative.
Quilty couldn’t possibly foresee every factor that might impact the way a film is received.
Horsman and Wood are also resolute in their belief that Quilty can accurately determine how a project “addresses the cultural moment” and give reliable box office projections. They pointed to *Revenge of the Nerds *as an example of a popular older film that would receive a lower Quilty score specifically because of the way it tries to depict sexual assault in a comedic light — something that modern viewers would see as being in poor taste.
When I asked Horsman and Wood about why Quilty gave Christy (which ultimately grossed around $2 million) a higher score than Sinners (which grossed $370 million), they insisted that the platform’s judgment “boiled down to the fact that Sydney Sweeney is really, really popular.” They said that on paper, Sweeney’s star power coupled with the fact that biographical dramas about boxing are cheaper to produce than fantasy / action features like *Sinners* made *Christy *a safer bet. But that situation highlights how Quilty’s logic isn’t all that reliable. Horsman and Wood admitted that there are some situations where Quilty couldn’t possibly foresee factors that might impact a film’s financial performance or the way audiences receive it.
Quilty could not, for example, have anticipated that Elijah Bynum’s Magazine Dreams (which Horsman produced) would end up being derailed by actor Jonathan Majors’ high-profile fall from grace in 2023. Similarly, nothing about *A Minecraft Movie*’s script would have indicated to the platform that the Chicken Jockey phenomenon would become part of the film’s monster success. Horsman and Wood told me that, eventually, they want Quilty to be able to see these types of things coming, but it’s hard to imagine how that might come to fruition.
For all of its fanfare, what Quilty is selling is roundabout access to an assortment of large language models that are being asked to predict the future as it relates to unproduced pieces of art. It would truly be amazing if any of these AI tools worked like Quilty claims they can. But most of them are just sophisticated pattern recognition / mimicry machines that are a long way out from being able to understand what humans find entertaining.
Follow topics and authors from this story to see more like this in your personalized homepage feed and to receive email updates.
- Charles Pulliam-Moore
-
-
-
-
関連記事
今日のまとめ
AI日報で今日の重要ニュースをまとめ読み