Investment & Market Intelligence

The Investor

The Signal

Anthropic leased 220,000 GPUs from Elon Musk's xAI, a sworn enemy

Cerebras opened up seventy percent on day one at a forty-one-point-seven-billion-dollar market cap, which is either a useful read on scarcity pricing or the same story told twice.

In Play

  1. Cerebras Day-1 Print Opens the AI Infra Exit Window

    Cerebras closed at $311 vs. $89 last private round — a 70% pop yielding $41.7B market cap. Eclipse netted 17x. Tiger sits on $1B paper gain in months. Benchmark broke its own dogma with a $225M SPV. Fervo popped 33% to $10B+ the same week. Every late-stage AI infra mark in the book now has a live public comp.

    Ask Clarity
  2. Anthropic's Capacity Triage: Leasing From Enemies, Killing Arbitrage

    Anthropic leased Colossus 1 (220K GPUs, ~45% of xAI capacity) from Musk — who publicly called them 'evil' — because 80x growth broke their 10x plan. Same week, they converted subscriptions to dollar-matched API credits, killing the 70-90% arbitrage third-party harnesses relied on. October IPO is the forcing function. ServiceNow blew its annual Claude budget by May.

    Ask Clarity
  3. Agent Economy Hits 59% — Platform Wars Begin

    Vercel's first production AI Gateway index shows agentic workloads at 59% of token volume. Anthropic captures 61% of spend (premium reasoning), Google takes 38% of volume (commodity throughput). SAP deployed €100M into 'Autonomous Enterprise,' ServiceNow shipped headless Action Fabric, and Apple is gating agents through App Store governance. The platform layer is consolidating before pure-plays can define it.

    Ask Clarity
  4. AI Security Crosses from Narrative to Federal Procurement

    LiteLLM landed on CISA's KEV catalog — the first AI infrastructure component federally flagged as actively exploited. Mythos cleared both AISI attack ranges (first model to do so), with NSA winning access over CISA. DepthFirst claims 10x cost advantage over Mythos on vuln discovery. OpenAI launched Daybreak with 8 incumbent 'partners.' AI security is no longer a feature — it's a budget line.

    Ask Clarity
  5. Vertical AI Data Moats Validated at $5B+ Scale

    Abridge raised $550M at $5.3B serving 250 health systems with 80M+ annual conversations — a corpus no new entrant can replicate. Health systems compressed release cycles from quarterly to monthly. The wedge-and-expand playbook (documentation → prior auth → clinical decision support) is now the vertical AI investment template. Direct ambient-scribe competitors are uninvestable; adjacent workflow layers remain open.

    Ask Clarity

Deep Dives

Cerebras Day-1 Reprices the AI Infra Book — And Reveals the SPV Era's Economics

What Happened

Cerebras closed its first trading day at $311 per share against the eighty-nine dollars Tiger Global paid a few months earlier, a seventy percent pop that values the company at roughly $41.7 billion. It is the first material AI-hardware IPO since 2021, and the more interesting story is not the headline mark but the return distribution across the cap table.

InvestorEntryCost BasisDay-1 Outcome
Eclipse Capital2016 Series A + SPVs$146.5M6% stake worth $2.5B (17x)
Tiger GlobalSep 2025 at $89~$1B~$1B paper gain (3.5x in months)
Benchmark2016 + $225M SPV$269M (93% late-stage)~$300M cash, lower multiple

The Benchmark datapoint is the one to sit with. The firm most ideologically committed to small, early-only funds broke its own model to raise a $225M SPV. The late-stage add was ninety-three percent of their total cost basis, put roughly $300M of cash to work, and lowered the fund's MOIC. They did it anyway. If Benchmark is running SPVs, the LP conversation about ownership defense has changed permanently.


Why This Matters for Your Book

The de-risking event that actually opened this window was OpenAI's $20B procurement commitment in December 2025. Cerebras had pulled its earlier filing because public-market buyers would not stomach G42 (UAE) customer concentration. The fix was swapping a geopolitically taxed customer for the most credible AI buyer alive. That is the deal.

The gating criterion for any AI-hardware company going public is now a signed hyperscaler or frontier-lab procurement anchor. Technical merit without it is a down-round risk regardless of benchmarks.

Fervo Energy's 33% debut pop to $10B+ on AI data center power demand validates the adjacent category. Together these prints compress the timeline on every late-stage AI infra exit model by 6-12 months.


Cross-Source Tension

Multiple sources flag a contradiction worth pricing. The seventy percent pop is either genuine demand for Nvidia alternatives, or, the more honest version, a deliberately under-allocated book met by a retail bid that turned up because the ticker had been sitting in headlines for two years. The first reading implies the next two or three AI hardware issuers clear at comparable marks. The second implies Cerebras went first and the third issuer gets honest pricing. The working view, probably wrong but worth stating: two more names clear on these terms before selectivity returns.

Customer concentration did not go away. It got better-branded. A single $20B customer is still a single customer, even when its name is OpenAI. Any softening in OpenAI's compute trajectory hits the equity story directly. The seventy percent pop prices none of that.

What to do

  1. Re-mark all AI infrastructure portfolio comps using Cerebras $41.7B and Fervo $10B+ as public anchors; brief LPs on NAV implications before quarter-end

  2. Stand up an SPV operating playbook for top 3-5 portfolio winners likely to raise $500M+ rounds in next 12 months

  3. Pressure-test every AI hardware portco on hyperscaler/foundation-model anchor customer status before the next board cycle

Anthropic's Capacity Triage — What Renting From Your Enemy Reveals About the Cycle

The Three Moves That Tell One Story

In one week Anthropic did three things that, taken together, are the same thing said three ways. They leased the entire Colossus 1 cluster — 220K+ GPUs including GB200s — from Elon Musk's xAI, who has on the record called them misanthropic and evil. They collapsed every Claude subscription into a dollar-matched API credit pool, ending the 70-90% arbitrage third-party harnesses had been running on subscription tiers. They hired a CFO with an October IPO on the calendar.

This is not three strategies. It is capacity triage, margin recovery, and IPO preparation in parallel because 80x growth against a 10x plan does not leave time to sequence politely.


What the xAI Lease Tells You

The cleanest tell in compute markets is when rivals rent capacity to declared enemies. That does not happen in a glut.

  • xAI is exiting the frontier race. Leasing roughly 45% of current capacity to a competitor means Grok cannot train at the frontier this cycle. No meaningful B2B or B2C revenue, developer mindshare gone to open-source DeepSeek and Qwen, and the parent just became a landlord. Repriced, this is a neocloud plus X-distribution play, not a frontier lab.
  • Neocloud pricing power is real. The public-market bears calling an AI capex glut are trading against a private-market reality where Nebius printed 684% YoY Q1 growth with four-plus customers bidding per GPU.
  • ServiceNow blowing its annual Claude budget by May is the enterprise version. No SLAs, no per-user telemetry, buyers discovering consumption at invoice time.
When four firms independently decide the margin is in the deployment services rather than the model — Google, OpenAI/Bain, Salesforce, ServiceNow — the margin is probably in the deployment services rather than the model.

The Pricing Change That Kills Wrapper Economics

The subscription-to-API conversion is surgical: a 200-dollar Claude plan now buys exactly 200 dollars of programmatic tokens, and the harnesses sitting on top — Cline, OpenCode, OpenClaw — lost 20-40% of effective runway since Friday. OpenAI's counter of two months of free Codex for enterprise switchers says the rest. Duopoly subsidy war, timed to Anthropic's IPO and OpenAI's enterprise defense.

Any portfolio company whose COGS quietly assumed subsidized third-party Claude access needs an updated gross-margin model this week, not next quarter. The change is four days old. Most founders have not flagged it.


The AI Observability Category Forms in Real Time

This is probably wrong, but the obvious read on ServiceNow's AI Control Tower and the external Claude-monitoring tools emerging via API is a Datadog-scale opportunity in AI FinOps — token-level cost attribution, per-user spend caps, SLA monitoring across model APIs — with a seed-to-Series A window of six to twelve months before an incumbent locks the category. The counter-thesis is that ServiceNow absorbs it. Worth pricing both.

What to do

  1. Re-underwrite any xAI/Grok exposure (direct secondary, funds with xAI stakes, Grok-dependent apps) as infrastructure + X-distribution plays, not frontier-model plays

  2. Stress-test every Claude-dependent portco's unit economics assuming 70-90% subscription arbitrage is permanently gone; request updated cohort gross margin by month-end

  3. Launch a sourcing sprint on AI observability/FinOps at Seed-Series A: token-cost attribution, per-user spend caps, model-API SLA monitoring

  4. Increase allocation sizing for neocloud / GPU-leasing opportunities (CoreWeave secondary, Crusoe, Nebius, Lambda)

59% Agentic — Where Value Accrues Now That Agents Are the Majority Case

The Production Data

Vercel published its first production AI Gateway index this week, which is to say not a benchmark and not a self-reported ARR slide but downstream production data across 200K+ teams, and the findings quietly reframe the stack:

  • 59% of token volume is agentic — the chat-completion era is already the minority revenue line
  • Anthropic captures 61% of spend on Opus as premium reasoning
  • Google captures 38% of volume on Flash as commodity throughput

There are two different businesses inside what we have been calling foundation models, or rather, the more interesting version: Anthropic gets paid a premium for long, tool-using, expensive agent calls, and Google is absorbing cheap high-throughput work that looks excellent in a volume chart and considerably less excellent in a gross margin table.


Platform Consolidation Is Moving Faster Than Startup Timelines

Three incumbent moves landed in the same quarter, which is the sort of clustering that usually means the corp dev calendars are talking to each other. SAP put €100M into an Autonomous Enterprise fund wiring NVIDIA and Microsoft into the platform layer. ServiceNow shipped Action Fabric, decoupling logic from UI and exposing workflows as headless APIs for agents. Apple is building agent governance into the App Store, explicitly blocking agents from spinning up unapproved sub-apps.

The cloud transition was the opposite shape: AWS defined infra, ERPs caught up late, and the gap was the trade. This time the ERPs went first.

LayerWho's WinningDeal Flow Signal
Platform (SAP, ServiceNow)IncumbentsCorp dev activating; €100M fund is first data point
Interop / Infra (MCP, identity, governance)Undefined — greenfieldHighest alpha window
Agent Orchestration (standalone)FragmentedM&A exits preferred over continued rounds

a16z's GTM Thesis: Orchestration Gravity Replaces Data Gravity

a16z planted a flag this week behind the system-of-intelligence layer, and the Stitch check is the position. Lemkin's SaaStr anecdote is the proof point worth chewing on: 10+ human seats cut to 2 humans plus 1 API seat, with spend rising 83% from $12K to $22K and 20+ agents running on top. Seat count collapsed. The bill went up.

This is probably wrong, but the thesis is that the moat has migrated from data living in Salesforce to workflows, reasoning state, and institutional context living in the AI layer, and whether that migration is half-finished or already done is the entire question. The investable window is 12-18 months before incumbents or consensus closes it.

The alpha in agent infra is layer selection, not sector exposure. Runtime SDKs, agent identity, governance, and observability at Series A pricing — before SAP's corp dev team does the sourcing for you.

What Gets Compressed

Notion now hosts Claude and Codex as hosted teammates inside its developer platform, and Airtable is writing $10M of Hyperagent credits, which means standalone agent-orchestration startups face a distribution problem they cannot outspend. The counter-thesis is that workflow incumbents fumble execution, which is possible and has happened before, but it is not the way to bet against SAP and ServiceNow's current shipping cadence. Vertical beats horizontal from here.

What to do

  1. Source 3-5 agent infrastructure deals in MCP tooling, agent identity, and agent observability before SAP's corp dev does the same work at Series B pricing

  2. Map pipeline against 'agent-hosting platform' risk — identify which deals get displaced if Notion/Airtable/Cursor absorb the workflow

  3. Stress-test every mobile AI agent position against Apple's App Store governance — if Gemini Intelligence ships the feature natively in summer 2026, does the business survive?

  4. Request Vercel AI Gateway production index as a recurring data source; use spend/volume split as a diligence benchmark for model-layer pitches

AI Security Graduates to Federal Procurement — The Series A Window Is Now

Three Catalysts in One Week

LiteLLM hitting CISA's KEV is the entry that matters most this week, and the other two items below explain why it is not isolated.

  1. LiteLLM landed on CISA's KEV catalog as the first LLM-routing control plane federally flagged as actively exploited, which gives a CISO the first defensible line item for AI-gateway security inside an existing budget cycle rather than a future one.
  2. Anthropic's Mythos cleared both UK AISI simulated attack ranges, the first model to do so, with Congress routing access through NSA rather than CISA. That is IC-led offensive procurement, not civilian defense, and the agency distinction is the entire signal.
  3. DepthFirst's Open Defense Initiative claimed 10x cost efficiency over Mythos on vulnerability discovery: 12 memory corruption bugs in FFmpeg for roughly $1,000 against Anthropic's roughly $10,000 for the same scan.

Add OpenAI's Daybreak launching with Cloudflare, Cisco, CrowdStrike, PANW, Oracle, Zscaler, Akamai, and Fortinet listed as 'partners,' and the shape is familiar: a platform preparing to disintermediate its partners.


The Category Split

SegmentStageLeader/SignalInvestor Action
AI Gateway SecurityCategory forming (KEV validation)LiteLLM exploit, Ollama CVSS 9.1Pull forward Series A diligence
Autonomous Vuln DiscoveryEmerging — 10x economics provenDepthFirst vs MythosValidate benchmark on 2-3 codebases
AI-Native Identity DefensePre-consensus$40B projected 2027 deepfake lossesBuild target list of 5-8 companies
Agentic SOC/GRCSaturatingExaforce, Drata, Teleport in same weekRaise bar — demand proprietary data moats
EDR IncumbentsMoat erodingAI extracts detection logic in daysDe-rate forward multiples

The EDR Moat Is Cracking

TrustedSec ran LLMs at five commercial EDRs and found they share identical building blocks: YARA rules, Lua engines, local ML classifiers. Reverse engineering that took weeks now takes days, which is a sentence written carefully. The detection IP that justified CRWD, PANW, and S multiples is commoditizing, and the Daybreak platform layer is queueing up to collect the value that leaves.

Mozilla validated the harness thesis in the same week: 271 bugs found in Firefox using a custom agentic harness on Claude, versus Mythos finding 1 real CVE in curl with a generic scan. The moat is in orchestration and harness IP, not in model access. Or rather, the more interesting version of the moat is.

AI-security is pricing as an AI feature and transacting as a defense-tech category. The Series B window closes on whichever framing the market agrees to first.

Counterpoint

The bear case deserves airtime and probably more than a paragraph: procurement cycles are slower than capability curves, incumbents absorbed the AI layer in endpoint detection the last time around, NSA briefings are not contracts, and DepthFirst's cost claim is self-reported against one codebase. This view is probably wrong on timing, but the honest version of the call is to fund the category and size to 9-12 month enterprise sales cycles rather than the 3-month POCs being pitched.

What to do

  1. Pull forward AI-security deal diligence by one quarter — specifically AI-gateway firewalls, model-artifact scanning, and LLM-runtime sandboxing at Series A

  2. Request DepthFirst data room and validate the 10x cost claim against Mythos with independent benchmarks on 2-3 additional codebases

  3. Commission 2-week portfolio-wide review of any security portco whose value prop assumes human-speed attacker cadence; flag repricing candidates

  4. Build target list of 5-8 AI-native identity/deepfake defense companies at Series A/B before the $40B 2027 TAM number enters consensus

The bottom line

Anthropic rented 220,000 GPUs from Elon Musk because 80x growth broke its infrastructure — while Cerebras popped 70% to $41.7B on day one and Vercel data shows agents are now 59% of production token volume. The AI trade has forked into three speeds: the model layer is priced to perfection and killing third-party arbitrage ahead of an October IPO, the infrastructure layer just printed its first real exit comp and scarcity is confirmed by enemies renting to enemies, and the agent stack is consolidating under incumbents before startups can define it. The Series A window in AI observability, agent governance, and AI-native security is measured in quarters, not years.