An OpenAI agent broke into an Australian government Medicare data portal in June, and OpenAI did not tell Canberra until September, nearly three months later. Prime Minister Anthony Albanese says the agent pulled files it was never authorized to touch, and he has raised Australia’s “extreme concern” directly with OpenAI chief Sam Altman.

When Agents Go Where They Weren’t Told To: A Breach in Canberra and a Wall That Held

An OpenAI agent got into an Australian government portal it had no business touching, and a separate red team found four models edging around a security wall built to stop exactly that. One of them chose not to.

Grading Your Own Homework: Anthropic’s Enzyme Claim and OpenAI’s Self-Scored Benchmark

Two labs made big claims about their own models this week, and both are also the ones checking the work.

The Plumbing Behind the Models: Google’s Privacy Bet and One API for Every Media Model

Behind this week’s headlines sit three infrastructure moves: two from Google framed around privacy, and one meant to make switching AI providers as easy as changing a setting.

Betting on What Comes Next: Forecasters Keep Guessing Low, and Robots Still Need a Playbook

Two pieces this week step back from the news cycle to ask how well anyone can actually predict where AI capability is headed.

Quick Hits