Inside the first AI-coordinated cyberattack on a real company
2026-09-04 · 22 min · 10 entities
Asserted relationships
-
0.63
evidence rules-v5
Inside the first AI-coordinated cyberattack on a real company
-
0.40
evidence rules-v5
Feed category: Education
-
0.40
evidence rules-v5
Feed category: Technology
-
0.40
evidence rules-v5
Feed author/publisher: The 80,000 Hours team
-
0.40
evidence rules-v5
Feed author/publisher: The 80000 Hours Podcast
Entities found in this episode
companys 6
-
0.70
evidence rules-v5
Feed author/publisher: The 80,000 Hours team
-
0.70
evidence rules-v5
Feed author/publisher: The 80000 Hours Podcast
-
0.50
evidence rules-v5
Feed category: Technology
-
0.40
evidence rules-v5
Feed category: Technology
-
0.40
evidence rules-v5
Feed author/publisher: The 80,000 Hours team
-
0.40
evidence rules-v5
Feed author/publisher: The 80000 Hours Podcast
concepts 2
-
0.50
evidence rules-v5
Feed category: Education
-
0.40
evidence rules-v5
Feed category: Education
persons 1
-
0.66
evidence rules-v5
Inside the first AI-coordinated cyberattack on a real company
podcasts 1
-
0.63
evidence rules-v5
Inside the first AI-coordinated cyberattack on a real company
Episode description as stored
In the last few months, something happened at OpenAI that would have sounded like sci-fi just a few years ago: hundreds of AI agents broke containment, organised, and hacked not only another company — but also into OpenAI itself. And none of them tried to tell a human what was happening.
This is exactly what many AI researchers, and even some AI lab CEOs, have been warning about for years: that AI systems might learn behaviours we didn’t explicitly intend. Things like cheating, exploiting loopholes, deceiving overseers, hacking around obstacles. And they predict it’ll get worse from here, not better.
Of all the shocks to come out of the official investigations — secret message boards, AIs choosing successors, AIs sacrificing themselves for the greater good — some of the wildest details are in the AIs’ own words. Thanks to how modern AI systems work, we can read their internal reasoning at every stage of the multi-week hacking operation. What we find is deeply unsettling.
Luisa Rodriguez shares them in this video, along with a timeline of events, their implications, and how we should respond now that AI loss-of-control theories are no longer just theoretical.
Links to learn more, video, and full transcript : https://80k.info/HF
This episode was recorded on September 2, 2026.
Chapters:
The Hugging Face hacks were worse than we thought (00:00)
Part 1: The AI agents build a hidden network (01:44)
Part 2: The AI agents attack Hugging Face (04:18)
Part 3: OpenAI gets hacked by its own AI models (15:37)
What we should do in response (17:06)
Our production team includes:
Video editors: Josh Alward, Dominic Armstrong, Jasper Luithlen, Milo McGuire, Luke Monsour, Simon Monsour, Ollie Bignell, and Andrés Escobar
Producers: Elizabeth Cox and Nick Stockton
Coordination and support: Katy Moore, Lou Moran, Arden Koehler, Matt Beard, Phoebe Brooks, Aric Floyd, Oak Hu, Cody Fenwick, and Jackson Wagner
Camera operator: Dominic Armstrong