AI Changelog
7One issue per week: what changed and how it changes our recommendations. Including what we filtered out, and why.
Open Weights Topped Arena. Google Stalled. Anthropic Blinked.
Moonshot AI's Kimi K3 beat Fable 5 on Arena coding at a third the price — the third Chinese frontier model in 30 days — on the same day Google confirmed Gemini 3.5 Pro's fourth delay since May. Anthropic ended Fable 5's four-week access crisis with a permanent subscription tier, and China launched a 29-nation parallel AI governance body.
The Gate Opened, Apple Sued, and Fable 5's Credit Cliff Slid Again
The government gate on frontier AI lasted 13 days: GPT-5.6 Sol reached GA as the first model to survive gating, and UK AISI confirmed Sol and Fable 5 share identical cyber vulnerabilities — making the original Fable 5 ban look like competitive weaponry. Apple sued OpenAI for trade secret theft, Meta entered the coding-model race with Muse Spark 1.1, and Microsoft began replacing OpenAI and Anthropic models in Office with in-house MAI.
The bans broke, the tokens multiplied, and the government cashed in
One week, three new government-AI relationship architectures: Fable 5 returned from its 17-day exile as the first government-ordered model shutdown ended, and Claude Sonnet 5 launched with near-Opus performance at Sonnet pricing — but a new tokenizer silently adds 30% more tokens, the year's first stealth price hike. OpenAI proposed handing the US government 5% of its equity, a $42.6B handshake, while California signed a half-price state deal with Anthropic, explicitly bypassing Washington.
The week defense AI went live
OpenAI's GPT-5.5-Cyber found 34 FreeBSD CVEs, 8 Linux kernel PoCs, and 5 Chrome V8 bugs in 5 days — the first frontier model deliberately deployed as a defense weapon at internet scale. The Five Eyes alliance warned the same day that AI cyberattacks are months, not years, away.
Price cuts did what benchmarks couldn't
Google cut Gemini 2.5 Pro pricing by 40% and flipped our budget long-context verdict without shipping a single new capability. Meanwhile the Mistral runtime verdict we promised last issue is in, n8n 2.0 will break your community nodes if you let it, and Zapier's agents went GA with pricing nobody can compute yet.
The week the coding verdict flipped
We changed our agentic coding pick for the first time in four months, OpenAI set a hard shutdown date you can't ignore, and Google quietly matched a capability everyone said was a moat. Nine things mattered this week — two change how you should work.
A million tokens changes the math, not the magic
Claude 4.6 Opus shipped a 1M-token context window and everyone declared retrieval dead. It isn't — but our long-document verdict did split for the first time, and the deciding axis is now corpus size, not model capability.
Get next week's issue
Every issue in your inbox the moment it's published. Nothing else, ever.