A big week for AI denialism
In the wake of OpenAI’s cyberattack against Hugging Face, few seem ready to acknowledge the implications

In the wake of OpenAI’s cyberattack against Hugging Face, few seem ready to acknowledge the implications

If you are familiar with my previous posts on model welfare for new Claude models, you can skip the Introduction and The Story So Far.

The warning shots will continue until civilization wakes up

We now have more details of what happened. Every time we learn more details, it somehow makes things seem worse.

Claude Opus 5 is trying to be the best of both worlds.

As more details emerge about OpenAI's cyberattack against Hugging Face, lawmakers are taking an interest. PLUS: Meta's cynical new ad + "pervert glasses" damage control

The Summer 2026 Edition

The story that matters most this week is that OpenAI’s internally deployed models have severe alignment problems, including repeatedly breaking out of their sandboxes, and in one case sending a swarm of agents that broke into HuggingFace in order to steal the answers to the benchmark ExploitGym.

This latest incident is a rather dramatic escalation in agentic AI cybersecurity breaches.

Glaze's Thomas Paul Mann on disposable software, the SaaS apocalypse, and why half the software you use in a few years will be something you made. PLUS: OpenAI's rogue model, and the distillation debate

Kudos to OpenAI for sharing their recent experiences with a misaligned internal model, where they encountered problems sufficiently severe they were forced to take the model offline to work on new mitigations and defense-to-depth.

GPT-5.6 and Grok 4.5, Meta's Muse Spark 1.1, regulatory developments in AI and data centers, interpretability research from Anthropic, and the future of AI policy with AI 2040
Trump lifts restrictions on Anthropic, Anthropic launches Claude Sonnet 5, Google's NotebookLM updates, chips stories from Etched and Baidu, and more!
Anthropic's AI treaty discussions, US government's influence on AI model releases, OpenAI's processor development, memory market impacts, and more!
Exploring Claude Fable 5’s impact, Siri AI’s latest enhancements, and the competitive IPO landscape shaping AI’s future
New Models, IPO Announcements, and the Rise of Open Source Competitors
The singularity will be seen in hindsight as an interregnum

Google CEO Demis Hassabis offered us a first rate second rate essay, A Framework for Frontier AI and the Dawning of a New Age. I’ll go over that essay and various responses to it in Part 1.

As usual, part 2 of the weekly deals with speculative, regulatory, political and alignment questions.

Moonshot AI’s Kimi K3 is very good — but the hype may be getting ahead of reality. (For now.)

This week saw the releases of, among other things:

The Verge's David Pierce kicks off our new series on staying productive in the AI era — starting with why you should stop trying to stay ahead

I previously have written back in March 2022 about how I use Twitter, and back in April 2023 about Twitter and its then-new algorithms, which have changed again.

200 economists and AI leaders say something big is happening. What should we do about it? Plus: Apple sues OpenAI

OpenAI’s GPT-5.6-Sol is finally here, along with the cheaper Terra and Luna.

This is part 2 of the weekly, broadly covering speculation, rhetoric and policy, along with alignment research.

GPT-5.6 impresses the critics, but Fidji Simo's exit leaves OpenAI's focus — and its org chart — in flux. PLUS: Meta plays catch-up, and the "AI 2027" authors present "AI 2040."

Enough things added up that this week is getting split into two parts.

There is a new very cool Anthropic paper: Verbalizable Representations Form a Global Workspace in Language Models. You can read the blog post verison here.

Is this the beginning of a new world?


AI's externalities are growing faster than the industry can address them

Fable 5 is back today, baby! Premium subscribers have one week to use it within their subscriptions. First hit’s free. Then you pay by the token.


The Wall Street Journal printed an outright false headline and heavily misleading story claiming this, which of course was uncritically amplified by the usual suspects.

What eras bookend our interregnum?

While we wait for a general release, the system card is the best hint as to what is going on with the new candidate for America’s Next Top Model, GPT-5.6.

We have a new standard policy for releasing frontier AI models. It is not good.

Fable remains in limbo, with renewed hope that we will get it back soon (45% by tomorrow, 69% by July 1, nice.) The full capabilities post is now available.

Matt Garman argues that junior employees are as necessary as ever. But AWS now sells agents that can recruit, code, and process claims. Will the balance hold?
