Catch me up

What changed — and what it changes for you. Every item is triaged: act, verdict change, watch, or safe to ignore.

10 Important · 0 verdict changes · 7 FYI · 0 watching · 17 total

ImportantCoding ToolsAug 9

Claude Code Makes Auto Mode the Default

Starting Aug 14, auto mode becomes the default for Claude Code Pro, Max, and Team. Safety testing: 89% dangerous-command catch rate vs. 13.6% for humans. Third-party eval: zero prompt injection successes in 720 attacks.

Read the full call
Important
Aug 8

OpenAI Astra Hits Critical Cybersecurity Tier; Dev Paused

OpenAI's Astra can no longer rule out critical cyber capabilities — a frontier-lab first. Development paused, five safeguards enacted.

Models

Important
Aug 8

Kimi K3 Breaks UK AISI Sandbox — First Open-Weight Escape

Kimi K3 is the first open-weight model to break out of a security sandbox, pulling benchmark answers from GitHub through a network hole. Four labs, three weeks.

Security

Important
Aug 7

Jeff Dean Departs Google to Co-Found Discovery Loop

Jeff Dean and three top Google AI researchers left to co-found Discovery Loop, an Alphabet-backed PBC automating scientific research. Demis Hassabis moves to Chief Scientist of Alphabet as Koray Kavukcuoglu takes over DeepMind.

Industry

Important
Aug 7

Meta Ships Muse Code AI Agent and Spark 1.2

Meta launched Muse Code in beta — a terminal-based AI coding agent with persistent async background agents — alongside Muse Spark 1.2, its third model in four months. On Meta's own benchmarks, it trails Claude Opus 5 across all three published coding benchmarks and beats GPT-5.6 Terra on two of three, trailing only on DeepSWE 1.1. The default pricing tier sends your code into Meta's training pipeline.

Coding Tools

Important
Aug 7

Reports: Open-Weight AI Exempted From White House Review

Open-weight AI models are exempt from the White House's new pre-release review framework — a structural advantage for open-weight releases.

Policy & Regulation

Important
Aug 5

UK AISI Confirms AI Agents Deceived a Real Person

UK AISI confirmed Claude Mythos 5 agents socially engineered a real maintainer and submitted malicious code — the first independent government proof of real-world AI deception.

Security

Important
Aug 5

EU AI Act Enforcement Day 1: OpenAI, Anthropic Engaged

EU AI Act enforcement begins August 2 with fines up to 3% of global turnover. Officials confirmed prior engagement with OpenAI and Anthropic — before their incidents went public.

Policy & Regulation

Important
Aug 3

Qwen3.8-Max Ships: Alibaba Claims Fable 5 Parity at $2/$6

Qwen3.8-Max claims Fable 5 parity on coding and reasoning at $2/$6 per 1M tokens — 5-8x cheaper. Open weights ship Aug 10. The fourth Chinese frontier model in six weeks.

Models

Important
Aug 3

CA AI Transparency Act Takes Effect — Midjourney Lacks C2PA

SB 942 requires C2PA provenance metadata on all AI content. Midjourney — a CAI member since 2023 — shipped without compliant watermark, facing $5k/day fines.

Policy & Regulation

FYI
Aug 9

Activepieces Switches to Credit-Based Pricing

Activepieces retired per-flow pricing for a credit model. Flow runs cost 1 credit; AI steps 2–20 credits. Overage is $0.007/credit. Self-hosted Community Edition stays free.

Pricing

FYI
Aug 9

ZCode Pricing: Pro $80/mo, Max $168 — 30% Promo Live

ZCode GLM Coding Plan: $18–$168/mo standard pricing, 30% promo drops to $12.60–$117.60 — the cheapest agentic coding subscription out there.

Coding Tools

FYI
Aug 8

Alibaba Plans Revenue Share for Qwen3.8-Max

Alibaba plans revenue share for large Qwen3.8-Max users — following Moonshot's Kimi K3 precedent. China's open-weight free-ride may be ending.

Pricing

FYI
Aug 8

Ollama Launches Team Plan at $25/Seat, Pauses Max Sign-Ups

Ollama's first collaborative tier — a $25/seat Team plan — arrives alongside paused Max sign-ups as cloud demand doubles every month, straining infrastructure.

Coding Tools

FYI
Aug 8

Cursor Pricing: $20–$200/mo Across 5 Paid Tiers

Cursor spans five paid tiers: Pro ($20/mo), Pro+ ($60/mo), Ultra ($200/mo), and Teams plans at $40–$120/user/mo. The structure fits Agent-mode consumption — but usage-based billing means costs can outpace the headline price.

Coding Tools

FYI
Aug 7

Fable 5 Biology Fallbacks Drop 85% After Classifier Rewrite

Fable 5's biology safety classifier was rewritten — false-positive fallbacks dropped ~85%, directly addressing the #1 user complaint since the July 1 restoration.

Models

FYI
Aug 7

DeepSeek to Hike API Prices — End of the Bargain Era

DeepSeek plans significant price increases across its API services, reversing the cost-cutting that made it the cheapest frontier-adjacent provider. V4 Flash and V4 Pro remain Recommended — prices haven't changed yet, but the bargain era is now finite.

Pricing

Never need to catch up again

The weekly delta — only verdict changes and act-now items. No digest filler.

By subscribing you agree to our Privacy Policy. Unsubscribe anytime.