Base44 が Claude Fable 5 を採用
エンジニアリングの最前線で作業する Base44 が、最も困難な開発業務に Claude Fable 5 を導入し、その信頼性を示した事例が報告された。
キーポイント
難易度の高い業務への適用
Base44 はエンジニアリングの最前線で作業しており、最も困難な開発タスクに対して Claude Fable 5 を信頼して導入した。
実務における信頼性の確立
単なる実験段階ではなく、実際の生産環境や重要なプロジェクトにおいて、このモデルが安定して機能していることが示されている。
Claude Fable 5 の性能評価
複雑なエンジニアリング課題に対処する能力において、Base44 がこのモデルを他社製品や以前のバージョンよりも優先して選定した背景が示唆される。
重要な引用
Working at the frontier: Why Base44 trusts Claude Fable 5 with their most challenging engineering work
Base44 trusts Claude Fable 5 with their most challenging engineering work
影響分析・編集コメントを表示
影響分析
この記事は、AI モデルが単なる実験段階を超え、企業の最も重要なエンジニアリング業務の核として信頼されるに至ったことを示す象徴的な事例です。Base44 のような先端的な組織が特定のモデル(Claude Fable 5)を選定したことは、その技術的優位性と実用性の裏付けとなり、業界全体における AI 導入の成熟度を高めています。
編集コメント
「Claude Fable 5」という具体的なモデル名が言及されており、AI モデルの進化が実務レベルで急速に浸透している様子が伺えます。Base44 が最前線の難題にこのモデルを投じる姿勢は、企業における AI 活用が「補助」から「中核的パートナー」へと役割を変化させた証左と言えるでしょう。
Base44 は、技術的なスキルに関係なく、誰でもフルスタックのアプリケーションやウェブサイトを作成できる「バイブコーディング」プラットフォームです。利用者は、開発者を抱えていない小規模企業から、SaaS プロダクトの開発に同社を活用している大企業まで多岐にわたります。
Base44 の創設メンバーであり、現在はプロダクト部門を率いる Yoav Orlev 氏にとって、仕事のやりがいを感じる最大の瞬間の一つは、時間や予算、あるいはノウハウが不足していた小規模企業が、このプラットフォームを使って何を実現できるかを目の当たりにすることです。デジタル店舗の構築から、飲食店のスタッフ向けシフト管理アプリの開発まで、その可能性は広範囲に及びます。彼のチームのミッションは、製品の機能を不断に拡張しつつも、誰もが使い続けられるように保つことにあります。
Base44 のプロダクトおよびエンジニアリングチームは常に迅速に動き、特に小規模から中規模のスコープを持つ機能の実装においては顕著です。しかし、相互依存する複数の部分に影響を及ぼすプラットフォームのコア部分への改変については、最も経験豊富なエンジニアにしか任せることができませんでした。
その典型的なボトルネックの一つが、システムプロンプトとその数百のバリエーションでした。これは、ユーザーが初めてアプリを作成するのか 5 つ目なのか、無料利用者とサブスクライバーのどちらか、そして構築中のアプリのカテゴリーや機能によって細かく異なります。もう一つの課題は、ネイティブモバイルインフラの変更でした。これにはモバイル分野に特化したエンジニアでなければ対応できませんでした。
2025 年初めにリリースされた Base44 のアプリ生成エンジンを支えてきた過去の Claude モデルでは、この重要な作業を任せるにはまだ不安がありました。Orlev 氏によれば、モデルがエラーに直面して行き詰まると、既存のコード内に解決策が存在する可能性に気づかず、同じ箇所での修正を試行し続ける傾向があったといいます。
「次に何をすべきかの判断は極めて重要ですが、過去のモデルは多くの場合、私が見る限り非常に素朴なアプローチを取っていました」と Orlev 氏は語ります。
Orlev 氏によると、Claude Fable 5 は初めて、ソフトウェアがどのように構築されるかを理解しているかのように推論できるモデルとしてチームにテストされました。
最も複雑な製品開発とエンジニアリング業務への信頼
Base44 では、新しい Claude モデルをさまざまな種類のアプリで評価し、レイテンシ、コスト、ビルドエラーの発生率などを測定しています。また、Minecraft のクローンを作成するなどのテストを通じて、モデルがゲームの物理演算やメカニクスをどのように処理できるかも検証します。
Claude Fable 5 では、2 つの点が際立っていました。まず、タスク完了までのターン数が大幅に減少した点です。そして、最初のプロンプトからより完成度の高いアプリを構築し、以前のモデルが見過ごしていたエッジケースにも対応できた点です。
チームは、これまで最上級エンジニアにしか任せていなかったタスクに Claude Fable 5 を投入しました。それは Base44 のシステムプロンプトの再構築です。約1時間のやり取りの後、Claude Fable 5 は約4時間にわたり自律的に作業を続け、必要な内容の90%から95%を完成させて返してきました。A/B テストインフラを活用したチームは、その日の午後にはこれらの変更を計測し、本番環境へリリースすることができました。
Claude Fable 5 が動作している間、Base44 の独自評価(evals)に抜け穴があることも発見しました。プロンプトの変更がキャッシュのヒット率に影響を与える可能性があり、数百万人のユーザー規模でそれが発生すればコスト増につながるにもかかわらず、チームはキャッシュヒットに関するテストを行っていなかったのです。モデルが盲点を指摘し、自ら修正を加えた形です。
また、Base44 のアプリ内エージェントを支えるハーン(harness)の変更で Claude Fable 5 が行き詰まった際、コードベースの他の場所で同様の問題解決策が見つかったはずだと推論し、該当箇所を調査して修正を導き出しました。Orlev はこう語ります。「『おそらくどこかで既に解決されているから、そこを調べてみよう』という推論プロセスは、他モデルではあまり見られない特徴です」。
Orlev によれば、Claude Fable 5 との協働は上級エンジニアと働くのに似ています。若手エンジニアには手順を細かく指示し、常に確認が必要ですが、上級者であれば目標とその背景にある理由(なぜそうするのか)さえ伝えれば十分なのです。
こうした取り組みはエンジニアチームに限った話ではありません。あるプロダクトマネージャーが、ネイティブモバイルアプリの構築を Base44 社内で行えるようにしたいと考えた際、Claude Fable 5 に任せてみたところ、約 2 時間半で本番環境への移行に必要な機能の 90% が整った開発環境が完成しました。
Claude Fable 5 の登場以前は、こうした作業は Base44 を支えるトップエンジニア 3 名か、専門家の空きを待つ必要がありました。現在は、モデルがタスクを実行し、Orlev 率いるチームがそのコードを検証・テスト・承認した上でリリースする体制になっています。

Claude Fable 5 は、Base44 のプロダクト・エンジニアリング・デザインチームに自信を与え、Sugeragents プラットフォームのより野心的な部分の開発を可能にしています。
次のステップ
Claude モデルの能力が向上するにつれ、Base44 チームが目指すプラットフォームの目標も進化しています。同社は、単にアプリを構築するツールから、構築したものを管理・成長させる支援をするプラットフォームへと進化させることを目指しています。現在公開されている Base44 Superagents は、これらのアプリを中心にワークフローを実行します。
複雑なタスクでも Fable 5 を信頼できることが分かっているため、Orlev は現在はプロダクトマネージャーやデザイナーに対し、以前は何かを壊す恐れから手を出せなかったプラットフォームの一部も積極的に構築するよう促しています。
「Fable は、事業においてより大胆な決断を下すための自信を与えてくれました」とオルレフは語ります。「製品をこれまで我々が恐れていた領域や可能性へと、全く新しい段階へと押し上げているのです。」
Claude Fable 5 で早速始めましょう。
原文を表示
Base44 is a vibe-coding platform that allows anyone, regardless of technical ability, to build full stack applications and websites. Its customers range from small businesses with no developers to companies using it to build full SaaS products.
Yoav Orlev, who joined Base44 as its first employee and now runs product, says one of the most satisfying parts of his work is seeing what small businesses can do with the platform for which they otherwise lacked the time, budget, or knowhow, whether that’s building a digital storefront or a shift-management application for restaurant staff. His team’s mission is to keep widening their product’s capabilities while keeping it usable for everyone.
The Base44 product and engineering teams have always moved quickly, especially when shipping small or medium-scope features. But any changes to the platform’s core that touch multiple interdependent parts could only be entrusted to the most senior engineers.
One such bottleneck was Base44's system prompt and its hundreds of permutations, which vary by whether someone is on their first app or their fifth, a free user or a subscriber, and by the category and features of the app being built. Another was changing the native mobile infrastructure, which only engineers with mobile expertise could do.
Earlier Claude models, which have powered Base44’s app generation engine since it launched in early 2025, couldn't be trusted with that work, Orlev suggests. When a model got stuck on an error, for example, it would keep working the spot in front of it instead of recognizing the fix probably already existed elsewhere in the code and searching for it.
“The decision on what to do next is a crucial one and most of the time [earlier] models would take, I would say, a naive approach,” he says.
Claude Fable 5 was the first model the team tested that could reason as if it had an understanding of how software is built, Orlev says.
Trusting Fable 5 with the most complex product and engineering jobs
Base44 runs each new Claude model through evals across different app types, measuring latency, cost, and build errors. The team also runs tests like building a Minecraft clone to see how a model handles game physics and mechanics.
With Claude Fable 5, two things stood out: it finished tasks in far fewer turns, and it built more complete apps from the first prompt, including the edge cases that earlier models skipped.
So the team pointed it at a task they had previously reserved only for the most senior engineers: rebuilding the Base44 system prompt. After about an hour of back-and-forth questions, Claude Fable 5 ran on its own for four hours and returned 90% to 95% of what they needed. Using its A/B testing infrastructure, the team was then able to measure and ship these changes that afternoon. And while Claude Fable 5 worked, it even flagged a gap in Base44's own evals: the team wasn't testing for cache hits, even though a prompt change can break the cache, and at the scale of millions of users that drives up cost. The model raised a blind spot and corrected it.
When Claude Fable 5 got stuck on a change to the harness behind Base44's in-app agent, it reasoned that the same problem had probably been solved elsewhere in the codebase, went to investigate that part, and came back with the fix. "This reasoning of 'this probably has been solved somewhere else, so I should go there to investigate' is something we haven't seen so often in other models," Orlev says.
Orlev compares working with Claude Fable 5 to working with a senior engineer. While a junior engineer needs every step specified and constant checking, you only need to brief a senior one on the goal and the why.
This type of work extends beyond the engineering team, too. When a product manager wanted to bring native mobile app building inside Base44, he pointed Claude Fable 5 at the job and after roughly two and a half hours had a working environment that was about 90% of what the team needed to move to production.
Before Claude Fable 5, this type of work had to wait for Base44's top three engineers or a specialist to free up. Now, the model executes tasks while Orlev's team reviews, tests, and approves the code before shipping it.

What’s next
As Claude model capabilities advance, so do the Base44 team’s goals for the platform. The team aims to turn Base44 from a tool that builds apps into one that also helps people manage and grow what they've built. Base44 Superagents, now public, run workflows around those apps.
Knowing that they can trust Fable 5 with complex tasks, Orlev now encourages product managers and designers to build in parts of the platform they were previously not willing to touch for fear of breaking anything.
“Fable has given us the confidence to make bolder moves with the business,” Orlev says. “It’s bringing the product to a whole new area and possibilities that before that we were, I would say, scared to do.”
*Get started with *Claude Fable 5*.*
関連記事
今日のまとめ
AI日報で今日の重要ニュースをまとめ読み