Four throughlines run today, starting with who pays for AI’s inference boom. Cerebras’ new CS-4 pairs with AMD and AWS chips instead of replacing GPUs, Jane Street tested Etched’s chip before leading a $700 million round at a $21 billion valuation, and Nvidia increasingly finances the demand it supplies.

Follow the Money: The Inference Hardware Fight

Three stories, one question: who is actually paying for the chips that run inference. Cerebras concedes GPUs aren’t going anywhere, a trading firm turned a pilot into a $21 billion bet, and Nvidia bankrolls the customers buying its own chips.

The Restraint Playbook: Choosing Control Over Speed

Four separate actors chose to slow down, lock down, or defend against misuse this week rather than ship faster. The throughline is control, over training, over governance, and over what a stripped-down model is allowed to say.

Shrinking the Frontier: Big Models on Small Machines

Frontier-scale performance keeps landing on hardware that used to be too small for it. The tradeoffs between local and cloud inference, and between a flat price and a real bill, are getting measured instead of assumed.

Agents Grow Up: From Demo to Infrastructure

Warp, Harvey, and Liquid AI are all publishing what actually happens when agents run in production instead of a demo. The pattern is agents becoming infrastructure, with the failure modes, memory systems, and scaling problems that implies.

Quick Hits

The rest of what moved today, in one line each.