Product & Strategy
The Product Desk
AT&T's 56%-for-2% trade is the first exchange rate a buyer can hold your product to.
The mechanism is unglamorous. A LiteLLM router sends the hard tasks to frontier models and everything else to open weights, which covers 40% of 45 billion daily tokens today and is targeted at 60-70% with OpenAI and Anthropic spend held flat. Any feature you price on model quality now competes against that ratio, refreshed quarterly by finance.
In Play
Open-Weight Routing Gets A Published Price
AT&T told The Information it will hold OpenAI and Anthropic spending flat while running 40% of its 45 billion daily tokens on open weights, with a stated target of 60-70%. After adopting the LiteLLM router, its AI coding costs fell 56% while measured quality fell 2%. That is the first at-scale exchange rate an enterprise buyer can hold your product to, and the same buyer caps what each developer may spend on premium AI tooling.
Ask ClarityModel IP Became A Rentable Input
Nvidia agreed to pay $6 billion for a non-exclusive license to Poolside's model technology plus $1 billion in equity at a $12 billion pre-money valuation, per a leaked investor letter obtained by Newcomer. Non-exclusive means AMD, Google or AWS can license the same technology next quarter. Fortune's Term Sheet reports investors marking 10-20% of portfolios as vulnerable to foundation-model encroachment. Any feature whose edge is model quality is a two-quarter asset, not a strategy.
Ask ClarityThe Harness Beats The Model
A research team wrapped database ACID semantics — validate before commit, isolate failed attempts, keep durable state — around a multi-step agent and beat Anthropic's Claude Code by up to 10.6% on a data-agent benchmark, with no better model. Two separate studies found memory-based self-improving agents look materially worse once task order and evaluation variance are controlled. Reliability is an architecture decision your backend team can ship today, and 'our agent learns from every session' will not survive a technical buyer's diligence.
Ask ClarityAgent Authorization Turns Into A Buyer Checklist
1Password, Cloudflare and Google each published a competing agent-identity architecture in the same week, and Uber open-sourced its Agent Detection & Response system with a public benchmark covering 300+ tasks, 17 attack techniques and 133 MCP servers. Independent vendors converging on one problem in one week is a spec forming, not thought leadership. Specs that form in August arrive as procurement questions in Q1, so your agent PRD needs an authorization section before a security reviewer writes one for you.
Ask ClarityAI ROI And AI Relief Split Apart
SolarWinds found 84% of IT service management teams say AI met or exceeded ROI expectations, while 52% report their total workload increased anyway. An SSRN working paper on Chinese secondary students found generative AI raised homework scores 18% while closed-book exam scores fell 20%. Census pulse data shows 55% of US workers used AI on the job, and 13% of those users report no time savings or more work. 'Saves X minutes per task' is a weak renewal story when the saved time gets reabsorbed by validation.
Ask Clarity
Deep Dives
- ●
AT&T Priced The Quality You Give Up For Cheaper Models
Routing stopped being a research spike: operator numbers, a flat-rate counterattack from Replit, and three published efficiency levers together reset what your margin defense has to contain.
What the 56% is actually measuring A developer inside Ask AT & T opens the model picker, sees GitHub Copilot, Devin, Claude Code and Codex, and takes whichever one she used last week. She is not optimizing anything. Her spending…
3 action items
- ●
Nvidia Rented A Frontier Model, Non-Exclusively, For $6B
Nearly the whole team that built the weights walked to the buyer, investors are marking down a tenth of their portfolios, and a 30-person publisher just demonstrated what a durable moat actually costs.
The mechanics behind the price Poolside had a six-week window at the end of 2025 to raise $2 billion for a 40,000 GB300 cluster. The round did not close and the cluster went away. Nvidia then hired 109 of the…
3 action items
- ●
Three Vendors Just Wrote The Agent Permission Spec For You
A one-click assistant exfiltration hole took roughly eight months to patch, an autonomous red-team agent beat an AI code reviewer, and enterprise buyers now have benchmark numbers to quote back at you.
Where the three architectures actually disagree They agree on the primitive: short-lived, scoped, logged grants that bind human plus agent plus intent. They split on where enforcement lives . 1Password keeps secrets away from the agent entirely; a trusted local…
3 action items
The edition continues
Take the signal into the room.
Sign up or log in to read all 3 deep dives in full, plus the final take.
Read the full editionContinue with LinkedIn