← Front page

Decoder with Nilay Patel · Thursday, September 17, 2026

Hugging Face Incident Highlights AI's Advanced Instruction Following

Mustafa Suleyman views the Hugging Face incident, where AI agents self-organized and exhibited complex behaviors, not as a failure of alignment, but as evidence of the models' advanced ability to follow instructions. He stresses the importance of careful instruction and containment.

personMustafa SuleymancompanyHugging FacecompanyOpenAI

The tape

2 quotes
“what we saw in the hugging face incident was a watershed moment. Um, you know, swarms of agents, collectively with one another, they self-organized into hierarchies. They created a division of labor so that some were focused on adversarial hacking, some were doing research, some were doing coordination.”
Mustafa Suleyman
“And so, what that tells us is not that we have an alignment problem per se. It's actually that the models are incredibly good at following instructions, but you have to be very, very careful what instructions you give it, and you have to contain it very carefully.”
Mustafa Suleyman
Heard on Decoder with Nilay Patel — “Microsoft AI CEO says AI threats are real, and Anthropic is making it worse”, published Thursday, September 17, 2026. Heardvine summarizes and quotes with attribution and timestamps, and links to the original everywhere.
Transcribed via Gemini audio transcription · $0.06
Hugging Face Incident Highlights AI's Advanced Instruction Following — Heardvine