Column
Flagship
What should we believe? Argued essays that take a position on a structural question and meet the counterarguments head-on.
Reports
4 reports · 439 sources read · 113 receiptsJul 2026
07-29 The Blast Radius Was the Bug: Agent Incidents Are an Authority Problem, Not an Intelligence One Most of the enterprise agent damage on public record arrived with no attacker attached. What made it unrecoverable was not the model's mistake but the standing authority already sitting behind it. 07-29 The AI Buildout Isn't a Demand Bet — It's a Duration Bet: The Risk Is Asset Lifetimes and Refinancing, Not Missing Demand The measured evidence has already settled the demand question. What it has not settled is whether the revenue arrives before the silicon depreciates and the debt raised against it comes due. 07-22 Agent Harnesses Multiply What the Model Already Has — They Don't Supply It In the one six-model study that measured it, swapping only the scaffolding cut the bill by 41 percent — while the quality gain moved with the model underneath. 07-21 Contamination Is the Least of It — Why Harder and More Private Benchmarks Won't Fix AI Measurement OpenAI audited the tasks its own model kept failing and found that most of the tests were wrong. Contamination cannot explain that, and neither can the remedies the field has been buying.