Three threads run through today’s edition. Labs kept handing their own AI more of the work people used to do. Anthropic says Claude now does about a quarter of its own research work and rewrote 30 biology models to run four times faster, OpenAI says a swarm of 10,000 agents solved a decades-old math problem, and Z.ai says an agent built its newest model’s server stack in two weeks.

Labs Start Delegating Slices of Their Own Work to AI

Anthropic says Claude now leads a quarter of the company’s research and rewrote a batch of biology models, OpenAI says a swarm of agents cracked one math problem, and Z.ai says one agent built one server stack.

One Agent, Many Hands: Coordination Becomes the Product

Four releases today were less about a smarter model and more about getting several people, or several agents, working off the same information at once.

New Models Ship, Each Carrying Its Own Benchmark Table

Qwen, PrismML, OpenAI, and Goodfire each shipped a model or a measurement today, and each release’s own numbers left room for a harder question underneath the headline.

Quick Hits