Most of this issue is about the distance between what AI is claimed to do and what anyone has shown. Fourteen of the fifteen stories carry an 8 October source date and one repository page carries none, so nothing here happened today.

Show Your Receipts: Big Claims, Thin Evidence

Four companies described what their AI can do. In each case the proof is missing, undated or limited to people who already have a seat.

Graded by the Author: Tests With a Catch

Three groups measured agents and published the result. Each one says plainly how the measuring was done, which is what makes the numbers worth reading.

Whose Number Is It: OpenAI Under Two Spotlights

Two OpenAI stories where the account depends on who is telling it and what they could see.

On the Builder’s Bench: Limits to Know Before You Ship

Four pieces for people running this stuff in production, each with a limit that is easy to miss on first read.

Today’s Quick Hits