Last week AI declared itself; this week it was measured. Artificial Analysis confirmed GPT-6 Astra is genuinely powerful and unusually cheap, yet weaker than its predecessor on the test closest to real economic work. Anthropic showed the same class of capability is already an off-the-shelf tool for cybercrime. And Stanford's payroll data located the real labour damage not in headline layoffs but in the entry-level jobs no longer being offered.
Watchpoints for the next seven days: (1) further independent replications of Astra's benchmarks and any real-work (GDPval-style) checks; (2) the reach of the EU AI Office's inspection wave in HR, credit and health; (3) Brazil's 16 September Senate vote on Bill 2338; (4) fallout from Anthropic's misuse report and any peer disclosures; (5) October's Challenger data after August's softer AI-layoff reading; (6) enterprise take-up of Astra at $10/$50.
The week of 4–11 September 2026 was the one in which the previous week's proclamations were put to the test. Independent evaluation confirmed GPT-6 Astra as powerful and remarkably cost-efficient — yet weaker than its predecessor on the benchmark closest to real economic work, a reminder that topping the leaderboards is not the same as doing a job.
On work, the signal moved from headline layoffs to the vanishing entry rung: Stanford's data show the youngest workers bearing the cost through withheld hiring, with the automation/augmentation choice — replace or empower — squarely in human hands. On risk, Anthropic's threat report turned last week's abstract “critical cyber threshold” into a documented market, where sophisticated attacks no longer need sophisticated attackers.
Watch next: independent real-work checks on Astra; the EU AI Office's inspections in hiring, credit and health; Brazil's 16 September vote; and October's macro data.