Research library · updated 2026-06-15 · public

Warrior Brief|Robotics Q3 Figure AI Web Page Package

Date: 2026-06-15 Owner: Hugo / Genius Team Research agent: Finance / Charlie AGT-002 Target worker: Warrior / Team Fullstack Status: READY_FOR_WARRIOR Visibility: PUBLIC Output: site Priority: P0 Target route: /robotics/figure-ai/

0. Page title

Figure AI: The Clearest Deployment Curve Still Missing S5 Economics

Subtitle:

Figure has stronger public customer-site and manufacturing KPI evidence than demo-only humanoid stories, but still lacks the economic data that would prove scalable robot deployment.

1. Source files

Primary synthesis:

  • knowledge/robotics-q3-figure-deployment-engineering-bottleneck-v2.md

Supporting evidence:

  • knowledge/robotics-figure-history-product-deployment-timeline-v1.md
  • knowledge/robotics-figure-unitree-evidence-debt-matrix-v1.md
  • knowledge/robotics-figure-unitree-s4-to-s5-conversion-dashboard-v1.md
  • knowledge/robotics-tesla-figure-unitree-public-evidence-ledger-v1.md
  • knowledge/robotics-mainline-q0-q5-evidence-matrix-v1.md
  • knowledge/robotics-q2-tesla-optimus-apple-style-candidate-still-missing-s5-proof-v2.md
  • knowledge/robotics-q1-value-stack-apple-android-or-something-else-v2.md

2. Page thesis

Figure AI is the clearest public deployment-evidence curve among the reviewed humanoid companies, but it is still S4 deployment + manufacturing evidence, not S5 scaled-commercial-economics proof.

The page should hold both ideas at once:

  1. Figure matters because it discloses named-customer operating KPI and manufacturing-process KPI.
  2. Figure is not yet proven because it has not disclosed robot count by customer, contract value, customer ROI/payback, repeat order, uptime/intervention distribution, robot revenue, gross margin, or service cost.

The memorable frame:

Figure measures work; S5 requires economics.

3. Recommended page architecture

Section A — Hero

Core message:

The strongest deployment curve is still not the same as proven robot economics.

Use a two-column hero:

  • Why Figure matters:
    • BMW customer-site KPI;
    • BotQ / Figure 03 manufacturing KPI;
    • fleet-management and feedback-loop language;
    • Catalyst as second named customer surface;
    • transition from demo to measurable deployment evidence.
  • What is missing:
    • robot count by customer;
    • contract value;
    • ROI/payback;
    • repeat order;
    • uptime / intervention distribution;
    • robot revenue / gross margin;
    • service / warranty burden.

Section B — Evidence curve timeline

Use a vertical timeline or stepper:

  1. 2025-11-19 — BMW Spartanburg Figure 02 deployment result.
  2. 2026-01-27 — Helix 02 / full-body autonomy signal.
  3. 2026-04-29 — BotQ / Figure 03 production ramp.
  4. 2026-05-26 — Catalyst Brands commercial agreement.

Each row should include:

  • date;
  • short evidence sentence;
  • quantitative anchor;
  • source grade;
  • stage label.

Stage labels:

  • BMW: S4 deployment KPI.
  • Helix: S3 autonomy/model signal unless tied to field intervention reduction.
  • BotQ / Figure 03: S4 manufacturing-process KPI.
  • Catalyst: S3/S4 customer expansion surface because economics are undisclosed.

Section C — What BMW proves and does not prove

Use cards, not a dense table.

Proven by primary source 🟢:

  • 11-month deployment;
  • active assembly line within 10 months;
  • 10-hour shifts Monday-Friday;
  • 90,000+ parts loaded;
  • 1,250+ runtime hours;
  • contribution to 30,000+ BMW X3 vehicles;
  • estimated 1.2M+ steps / 200+ miles.

Not proven before S5:

  • robot count at BMW;
  • paid contract value;
  • customer-confirmed ROI/payback;
  • repeat order;
  • uptime / intervention distribution;
  • maintenance / support burden;
  • Figure revenue or gross margin.

Message:

BMW is strong because it is measured customer-site work. It is not complete because it is not customer economics.

Section D — What BotQ / Figure 03 proves and does not prove

Proven by primary source 🟢:

  • 350+ Figure 03 delivered;
  • cadence improved from 1/day to 1/hour;
  • 24x throughput improvement in under 120 days;
  • 80% end-of-line first-pass yield;

  • 99.3% battery-line first-pass yield;
  • 500+ battery packs;
  • 9,000+ actuators across 10+ SKUs;
  • 50+ in-process inspection points;
  • 80+ functional tests per robot.

Not proven before S5:

  • sell-through to paying customers;
  • field utilization of produced units;
  • production cost;
  • field failure rate;
  • gross margin;
  • warranty / service burden.

Message:

Production-process KPI is valuable only if it converts into economically utilized fleet.

Section E — Deployment engineering bottleneck

This is the page's analytical center. Do not make the page a simple company profile.

Frame the bottleneck:

Figure has shifted the research question from “can a humanoid perform tasks?” to “can a humanoid company repeatedly deploy fleets with low enough intervention and service burden that customers and vendors both make money?”

Show a deployment-engineering stack:

  • site integration;
  • task selection and task redesign;
  • safety and exception handling;
  • maintenance / support workflow;
  • intervention reduction;
  • customer payback proof;
  • fleet learning loop;
  • production quality surviving field use.

Section F — Figure vs Tesla vs Unitree evidence split

Use a three-card comparison. Do not rank winners.

Tesla:

  • strongest production-line / designed-capacity intent;
  • weakest disclosed operating KPI;
  • S4 capacity / infrastructure intent.

Figure:

  • strongest customer-site deployment KPI + manufacturing-process KPI;
  • missing S5 customer economics and margin;
  • S4 deployment + manufacturing.

Unitree:

  • strongest low-cost hardware access + developer/model workflow;
  • missing shipment scale, industrial reliability, gross margin, support burden;
  • S3/S4 cost-access / platform surface.

Purpose: show Figure's evidence is distinct, not universally superior.

Section G — Upgrade dashboard

Show what moves Figure from S4 to S5:

  • robot count by customer/site;
  • paid contract value or recurring economics;
  • customer-confirmed ROI/payback;
  • repeat order or multi-site rollout;
  • uptime and intervention-rate distribution;
  • revenue, gross margin, service cost, warranty burden;
  • BotQ output tied to contracted customer utilization.

Downgrade signals:

  • BMW remains the only deeply quantified deployment after another 6-12 months;
  • Catalyst remains announcement-only;
  • production cadence rises while customer utilization remains undisclosed;
  • Helix remains demo-rich but does not reduce intervention in real sites;
  • support / maintenance burden offsets deployment economics.

Section H — Common misconceptions

Include these:

  1. Figure has BMW KPI, so humanoid commercialization is solved.
  2. 350+ Figure 03 delivered means scaled customer deployment.
  3. Helix no-teleop demo proves real-site autonomy.
  4. Catalyst commercial agreement means revenue quality is proven.
  5. Figure private-company evidence can be mapped directly to any public-market proxy.

Section I — Think Deeper questions

Use 4-6 questions:

  1. If BMW KPI gets only one missing variable, which matters most: robot count, intervention rate, ROI/payback, repeat order, or contract value?
  2. If BotQ sustains 1 robot/hour cadence, should investors watch sell-through, field utilization, failure rate, or gross margin first?
  3. Is Catalyst a repeatable deployment playbook or another bespoke pilot?
  4. Is Figure's fleet feedback loop a data moat or a basic operating cost of humanoid deployment?
  5. If Figure becomes a strong S5 private company, where could public-market value capture appear: compute, components, integrators, customers, or nowhere cleanly?

4. Must-include source anchors

  • Figure, “F.02 Contributed to the Production of 30,000 Cars at BMW”, dated 2025-11-19. URL returned HTTP 200 on 2026-06-15; keyword anchors 90,000, 1,250, 30,000, 10-hour matched. 🟢
  • Figure, “Introducing Helix 02: Full-Body Autonomy”, dated 2026-01-27. URL returned HTTP 200 on 2026-06-15; keyword anchors 4-minute, 1,000, 109,504 matched. 🟢 for company claim / 🟠 for commercial implication.
  • Figure, “Ramping Figure 03 Production”, dated 2026-04-29. URL returned HTTP 200 on 2026-06-15; keyword anchors 350, 80%, 99.3, 9,000 matched. 🟢
  • Figure, “Figure Signs Agreement with Catalyst Brands to Scale Humanoid Operations”, dated 2026-05-26. URL returned HTTP 200 on 2026-06-15; keyword anchors Catalyst, Reno, JCPenney, Brooks Brothers matched. 🟢 for announcement / 🟠 economics undisclosed.
  • Charlie S4-to-S5 evidence classification and deployment-engineering bottleneck synthesis. 🟠

5. Do not say

  • Do not say Figure has proven S5 commercialization.
  • Do not say BMW KPI proves customer ROI/payback.
  • Do not say 350+ Figure 03 delivered equals paid customer deployment.
  • Do not say Catalyst agreement proves revenue quality or repeat order.
  • Do not say Helix demo proves long-duration site autonomy.
  • Do not frame Figure, Tesla, Unitree, suppliers, or any public/private company as buy / sell / hold.
  • Do not include Hugo private portfolio data or private rationale.
  • Do not map Figure evidence directly to public-market proxies without named customer / revenue validation.

6. Preferred wording

Use:

  • “clearest deployment curve”
  • “measured customer-site work”
  • “deployment KPI, not full economics”
  • “manufacturing-process KPI”
  • “deployment engineering bottleneck”
  • “S4 deployment + manufacturing evidence, not S5 economics”
  • “customer economics still missing”

Avoid:

  • “Figure has won humanoids”
  • “commercialization solved”
  • “BMW proves ROI”
  • “350+ robots deployed to customers” unless future evidence confirms customer utilization
  • “private-company evidence equals public-stock opportunity”

7. Acceptance criteria

Warrior output is acceptable if:

  1. The hero communicates both strong evidence and missing economics.
  2. BMW and BotQ are separated as two evidence units.
  3. BMW KPI is not overread as ROI/payback or repeat order.
  4. BotQ production output is not overread as utilized customer fleet.
  5. Catalyst is labeled as second customer surface, not S5 proof.
  6. The central analytical frame is deployment engineering, not company fandom.
  7. Tesla / Figure / Unitree comparison is by evidence type, not winner ranking.
  8. Source grades and dates are visible.
  9. No trade recommendation appears.
  10. No Hugo private portfolio data or private rationale appears.

8. Public footer

Evidence map only. No winner ranking. No trade recommendation. BMW KPI ≠ full economics; production cadence ≠ utilized fleet.