Leadership & Executive

The Board Room

The Signal

Haiku 5.5 ties GPT-6 Luna on the price sheet and loses on the bill.

The five-point lead at maximum effort is real, and it costs roughly three times the rival's tokens to earn. The tokenizer then counts the same text 25-30% heavier, and prompts past 100,000 tokens pay five times the base rate. That means any cost model you ran on list price will hand the long-context jobs to the model that bills them worst.

In Play

  1. Small-Model Price War Hides in the Fine Print

    Anthropic priced Claude Haiku 5.5 at $0.10 per million input tokens and $0.50 per million output tokens, the same as OpenAI's GPT-6 Luna. Artificial Analysis scores Haiku higher, 43 to 38. Prompts over 100,000 tokens pay five times the base rate, and Haiku's tokenizer counts the same text as roughly 25-30% more tokens. If your cost model uses list price, it will pick the wrong vendor for long-context work.

  2. Frontier Lab Revenue Was Overstated

    The Financial Times puts OpenAI's annualized revenue near $50B at the end of September, not the widely reported $70B. Nvidia, CoreWeave and Nebius all fell on the report. Many vendor-viability models and capacity plans are anchored to lab revenue. That revenue is unaudited, estimated as one month times 12, and calculated differently by OpenAI and Anthropic.

  3. OpenAI's Outside Oversight Is Thinning

    OpenAI fired three safety researchers, including Tomek Korbak. Korbak was its main technical contact with the outside auditor METR while METR reviewed July's agent containment incident. The researchers say they were fired for putting safety first. OpenAI reportedly says they mishandled confidential information. Your vendor reviews probably score model safety, which still looks strong, but not whether outside auditors keep their access.

  4. Agents Become a Customer Class

    Agents generated 48% of usage of Cloudflare's Wrangler command-line tool in one week. Meta and Sierra launched the Personal Agent Protocol with Shopify, Stripe and Walmart as partners. Stripe and Cloudflare are each building ways to charge agents per request. Your seat-based pricing and employee-based identity model were designed for human buyers.

  5. AI Returns Accrue to Builders

    A third of The Information's tech-forward subscribers say AI returns a multiple of what they spend. A year ago, MIT found that only 5% of custom enterprise tools reached production. Separate research finds just 7% of companies have scaled AI, so the survey probably reflects organizations that build software. Respondents also report using Microsoft Copilot less even as seats grow, which suggests unused seats in your licensed tools.

Deep Dives

  1. Haiku 5.5's Price Sheet Is Not Its Bill

    Matching a rival's per-token price means little when a model reasons longer, counts text differently and charges five times as much past a context line.

    Where the discount leaks out Three mechanics sit between the headline price and your invoice. The first is token burn . Artificial Analysis found that Haiku 5.5 needs about 55,000 output tokens to match GPT-6 Luna's index score of 38.…

    3 action items

    ●
  2. Two Corrections to Your OpenAI Vendor File

    These developments weakened both the revenue figure behind vendor-viability models and the outside audit behind safety assurances, and most scorecards track neither.

    A run rate built from leaks The revenue gap comes from how the figure was calculated . The Information traces the overstatement to "clunky calculations" by investors and news outlets. OpenAI has restated nothing. On the usual convention of one…

    3 action items

    ●
  3. Agents Are Arriving as Customers Your Seat Pricing Can't Bill

    A commerce coalition, a payments race and tools rebuilt for machine traffic all lead to the same P&L question: what does an agent pay for?

    The load arrived before the business model GitHub had to rebuild its Git infrastructure because agent-driven development generates heavy concurrent reads and writes. After separating storage from compute, it reports up to 35x higher write throughput , a figure it…

    3 action items

    ●

The edition continues

Take the signal into the room.

Sign up or log in to read all 3 deep dives in full, plus the final take.

Read the full edition

Continue with LinkedIn