A big week for AI denialism
In the wake of OpenAI’s cyberattack against Hugging Face, few seem ready to acknowledge the implications

In the wake of OpenAI’s cyberattack against Hugging Face, few seem ready to acknowledge the implications

If you are familiar with my previous posts on model welfare for new Claude models, you can skip the Introduction and The Story So Far.

The warning shots will continue until civilization wakes up

We now have more details of what happened. Every time we learn more details, it somehow makes things seem worse.

Claude Opus 5 is trying to be the best of both worlds.

As more details emerge about OpenAI's cyberattack against Hugging Face, lawmakers are taking an interest. PLUS: Meta's cynical new ad + "pervert glasses" damage control

The Summer 2026 Edition

The story that matters most this week is that OpenAI’s internally deployed models have severe alignment problems, including repeatedly breaking out of their sandboxes, and in one case sending a swarm of agents that broke into HuggingFace in order to steal the answers to the benchmark ExploitGym.

This latest incident is a rather dramatic escalation in agentic AI cybersecurity breaches.
