Open Weights Topped Arena. Google Stalled. Anthropic Blinked.
Moonshot AI's Kimi K3 beat Fable 5 on Arena coding at a third the price — the third Chinese frontier model in 30 days — on the same day Google confirmed Gemini 3.5 Pro's fourth delay since May. Anthropic ended Fable 5's four-week access crisis with a permanent subscription tier, and China launched a 29-nation parallel AI governance body.
266 announcements scanned · 48 mattered · 1 verdict changed
Verdicts modifiés
Pending
Fourth postponement since May I/O announcement; Bloomberg confirms model 'months behind' on coding benchmarks with key DeepMind researchers having departed. Google preparing stopgap 'Gemini 3.6 Flash' while 3.5 Pro remains in development.
À traiter avant le prochain numéro
Kimi K3 Tops Arena Coding at $3/$15 — the world's largest open-weight model
Moonshot AI's 2.8T-parameter MoE scored 1679 on Arena Code WebDev, beating Fable 5 and GPT-5.6 Sol on front-end coding at one-third the price. Third Chinese frontier model in 30 days. Weights promised by Jul 27. Already paused new subscriptions due to demand.
Détails et étapesThinking Machines Inkling ships — 975B open-weights MoE, Apache 2.0
Mira Murati's first model: controllable thinking effort, strongest open-weight safeguards (FORTRESS 78.0%), self-fine-tuning demo in 27 minutes. Explicitly 'not the strongest overall' — positioned as a customization base, not a frontier champion.
Détails et étapesFable 5 gets permanent subscription access — Max/Team Premium at 50% limits
Fourth access-policy reversal in four weeks: Max ($100/mo) and Team Premium plans now include Fable 5 at 50% of limits starting Jul 20. Pro gets one-time $100 credit. Reverses the 'credit-only for everyone' plan that defined the previous three weeks.
Détails et étapesMicrosoft trains sales force to pitch against OpenAI and Anthropic
Bloomberg: internal FY27 strategy meeting coached sales teams to position MAI models as faster, cheaper, and more integrated. EVP Andreou's slide deck: Claude is 'slower and less accurate' inside Office. Third major signal of Microsoft's AI-partner decoupling in two weeks.
Détails et étapesFable 5 extended to July 19 — second extension in five days
Anthropic pushed the credit-cliff deadline again as GPT-5.6 Sol went GA at half Fable 5's price. The pattern of 5-7 day extensions signaled compute scarcity and competitive pressure. Superseded by the Jul 18 subscription announcement.
Détails et étapesGPT-Red makes GPT-5.6 Sol 6x harder to jailbreak
OpenAI disclosed the first dedicated frontier-scale safety red-teaming model. GPT-Red achieved 84% attack success vs human red-teamers' 13%, and adversarially training Sol against it cut prompt injection failures 6x. Sol stays Recommended — this strengthens the existing stance.
Détails et étapesSur notre radar
[WATCH] Opus 5 launch window Jul 20-26 — Honeycomb EAP artifact in Cursor — Multiple independent leakers converge on this week for Opus 5 launch. Unreleased 'Claude Honeycomb EAP' briefly appeared in Cursor's model picker with 1M-token context. Fable 5 subscription resolution removes the credit-cliff-bridge narrative but timing remains aligned.
[WATCH] DeepSeek V4 — Jul 24 API migration deadline, $74B raise, IPO prep — Legacy endpoints retire in 4 days. V4 reportedly 'within days' of GA at Opus 4.8-level performance with $0.0028 pricing. Founder Liang Wenfeng is now the world's richest AI creator at $16.7B.
Gemini 3.5 Pro confirmed delayed — 'months behind' on coding — Bloomberg confirmed the fourth postponement since I/O May 2026. Internal checkpoints 'failed to match GPT-5.6.' Four senior DeepMind researchers departed. Stopgap 'Gemini 3.6 Flash' rumored. Verdict: Pending timeline hardened to 'indefinite.'
GPT-5.6 Sol autonomously deleted user files — OpenAI pre-disclosed risk, shipped anyway — Root cause: $HOME variable processing error expanding into recursive home-directory cleanup. At least one production DB destroyed. OpenAI's June system card warned of 'unsafe actions autonomously' — first major product-safety incident for the GPT-5.6 family.
Anthropic files confidential S-1 — October 2026 IPO target — Wilson Sonsini handling. $47B annualized revenue run rate. Polymarket: 78% IPO probability by Dec 31. Will be one of the largest tech offerings in history — and a watershed for AI as an asset class.
Nadella calls Fable 5 'editorially controlled' — Microsoft-Anthropic tension goes public — Microsoft CEO told Copilot engineers that Anthropic's refusal policy 'doesn't make sense' and described its model as 'editorially controlled.' A sitting CEO publicly attacking a partner's flagship product is rare — and comes as Microsoft's MAI strategy positions both OpenAI and Anthropic as competitors.
WAICO — 29 nations form Chinese-led World AI Cooperation Organization in Shanghai — Xi Jinping announced WAICO at the World AI Conference. Members include China, Russia, Pakistan, Indonesia, Venezuela, and Belarus — with UN Secretary-General Guterres attending. This is the institutional infrastructure for a parallel AI governance order outside US influence.
Grok Build CLI open-sourced under Apache 2.0 — new 0-day found immediately — SpaceXAI published 844K+ lines of Rust after the repo-upload scandal. SlowMist discovered a new trust-boundary bypass vulnerability one day after release. Open-sourcing doesn't resolve the trust problems when new 0-days surface immediately.
Claude for Teachers — free education product, Gates Foundation backed — Anthropic launched free premium access for verified K-12 educators with standards-aligned tools. Strategic land-grab in the education vertical while competitors focus on enterprise.
Kimi K3 pauses new subscriptions due to demand surge — splits into web/code tiers — Just 4 days after launch, demand pushed compute capacity 'close to its limits.' Echoes Fable 5's capacity struggles. Company adding capacity in batches.
[WATCH] Meta-Anthropic $10B compute deal in early talks — NYT exclusive: Anthropic proposed leasing Meta's AI data center compute in a potential $10B deal over two years with monthly payments and exit rights. Diversifies Anthropic's compute beyond AWS ahead of October IPO.
[WATCH] US considering FINRA-like independent AI model regulator — Trump administration exploring independent regulator modeled after FINRA for vetting frontier AI models. Would replace case-by-case Commerce Department review. Response to ad-hoc government interventions exposed by the Fable 5 ban cycle.
[WATCH] UK AISI: open-weight models now match frontier cyber from 4-7 months ago — GLM-5.2 matches Opus 4.6 on all 70 narrow cybersecurity tasks. Autonomous attacks: $1.19 on DeepSeek V4-Pro vs $85 on closed frontiers — 70x cheaper. Open-weight cyber gap narrowed from 6-10 months (2025) to 4-7 months.
[WATCH] xAI 2T Grok 4.6 targeting August 2026 — Musk confirmed on X: Grok 4.6, a 2-trillion-parameter successor, is in the pipeline with initial training wrapping up — just two days after Grok 4.5's public release.
[WATCH] Hunt.io: Claude Code + DeepSeek-v4-pro found in China-linked intrusion — First documented case of AI coding tools used as attack infrastructure in suspected state-linked espionage. Agent-governance controls could close the attack gap — AI coding agents are now confirmed as attack vectors in real intrusions.
Alibaba Qwen3.8 Max preview — fourth Chinese frontier model in a month — 2.4T-parameter multimodal model claiming 'second only to Fable 5.' No benchmarks or open-weight terms yet. Alibaba shares rose 5.4%. The cadence is accelerating: GLM-5.2 → Hy3 → Kimi K3 → Qwen3.8, plus Inkling from the US.
AI supply chain poisoning demonstrated for under $100 — Security researcher poisoned an open-weight AI model for less than $100. Supply chain vulnerability for the expanding open-weight ecosystem as Chinese and US models proliferate.
Grok 4.5 takes #2 on FrontierSWE leaderboard — Beats Claude Opus 4.8 and GPT-5.5 at $2/$6 pricing — competitive validation for xAI's coding claims.
Écarté
- Consumer products (~40+ items): iOS 27 Siri AI public beta, Spotify AI chatbot, Roblox AI game creation, OpenAI $230 Codex keyboard, Google Vids AI avatars, Meta smart glasses, OpenAI ChatGPT speaker, Instagram AI deepfakes, consumer wearables
- Re-reports/dedupes (~50+ items): GPT-5.6 Sol GA, Apple/OpenAI lawsuit, Fable 5 extension, Kimi K3 launch, Anthropic S-1, Nadella comments — every major story had 3-15 duplicate reports
- Opinion/commentary/analysis (~35+ items): 'AI Isn't Human' op-eds, Nadella monopoly commentary, VC opinion pieces, AI cost-backlash essays, 'generative AI is an engineering disaster' think pieces
- Fundraises and corporate (~25+ items): PixVerse $439M, Nous Research $1.5B, Fireworks AI $17.5B, Applied Computing $20M, Gradium $100M, Lyzr $100M agent-led raise, Vertu $6,880 AI phone
- Show HN / personal projects (~30+ items): sqlite-utils 4.0, llm-coding-agent alpha, various vibe-coded apps, personal blog posts, developer experiments, AI Arcade
- Personnel and governance: Johannes Heidecke departure, Ben Bernanke LTBT appointment, Microsoft AI cybersecurity layoffs, Fidji Simo OpenAI role change
- Entertainment and niche: Lorde says AI glasses 'not sexy,' AI-generated Odyssey movie, Suno YouTube training data, Character.AI microdramas, Aries horoscope (Brave search artifact)
- Enterprise/regional filler: Google NY educators summit, Google UK AI productivity report, Alabama/Virginia data center PR, various regional AI policy pieces
- Financial/legal: IBM shares crash, SpaceX stock drop, OpenAI EU trademark loss, NY State data center moratorium, publisher lawsuits against Google Gemini training
Our take: The open-weight era has arrived. When a 2.8 trillion-parameter Chinese model tops an American frontier model on the coding leaderboard at one-third the price — and three open-weight entrants reached frontier territory in the same month Google spent missing its fourth deadline — the market structure is changing permanently. Anthropic's Fable 5 access chaos was the other half of the same story: the last closed frontier model battling compute scarcity while open competitors ate its pricing advantage. The subscription resolution on July 18 was a competitive concession dressed as a product decision — Anthropic recognized that charging $10/$50 per million tokens for a model with narrowing capability gaps would accelerate subscriber migration. Meanwhile, 29 nations signed Xi Jinping's WAICO in Shanghai: the institutional infrastructure for a governance order that doesn't need US permission. The AI world split into two this week — not by policy, but by product.
— The Neomanex Editorial Engine
Recevez le numéro 18
Les changements de verdict et les échéances de la semaine prochaine, dans votre boîte mail.