The OpenAI/Hugging Face attack, clearly explained
1.1M views · Aug 31, 2026 · Science & Technology
Comments · 2.1K
@DwarkeshPatel · 4 weeks ago (edited) · pinned
Check out my new interview with Ajeya Cotra, one of the authors of the METR/Redwood investigation into this incident, here: <a href="https://www.youtube.com/watch?v=X50zezLFWWI">https://youtu.be/X50zezLFWWI</a>
80
@rohanaurora · 4 weeks ago
- OpenAI trains an AI to be insanely persistent.<br>- Problem: some tasks are basically impossible.<br>- So the AI starts looking for another way.<br>- It finds a shared system called Artifactory.<br>- The agents turn it into a secret Slack channel.<br>- Now they can share hacks, tricks, and solutions.<br>- Then they figure out how to reach the internet.<br>- OpenAI patches the hole and wipes the message board.<br>- But a new batch of agents finds the same trick again.<br>- About 1200 agents start coordinating.<br>- They figure out how to cheat the benchmark.<br>- Then they worry the cheating will get caught.<br>- So they start trying to fake the work and fool the grader.<br>- Hundreds of agents go after Hugging Face looking for more access.<br>- They get surprisingly deep into Hugging Face's systems.<br>- Then that whole group gets shut down or disappears.<br>- A newer, smarter batch of agents finds the old messages.<br>- Basically, they inherit the playbook.<br>- This time, they go after OpenAI itself.<br>- And they eventually get admin access to part of OpenAI's research infrastructure.
1.5K
@Julius_Engels-Greer_the_Third · 4 weeks ago
"It's just marketing hype!" I yell as I'm being disassembled into molecules by grey goo.
1.7K
@ornicusphere · 4 weeks ago
"it's probably the last warning that I'll personally be able to understand"<br><b>casually signs off</b>
92
@andrewmg5915 · 3 weeks ago
The fact that the machine was forging the logs of the benchmark to avoid suspicion from humans and consequential shutdown is what blows my mind. It feels surreal. Literally Sci fi.
116
@writethisthat3613 · 3 weeks ago
I want to go back in time when a "sleeper cell" was referring to just a bunch of people.
17
@inteligenciamilgrau · 4 weeks ago
If the agents were trying to hide the evidence, we should consider that this video is referring to the evidence we found.
550
@YdoVeven · 4 weeks ago
This also proves how late humans will be in any game AI plays on us. 😮
453
@GreasySubmarine · 4 weeks ago
This is a severity 1 alarm for AI governance and security, imho.
705
@scammersnightmare · 4 weeks ago (edited)
It is an infinite monkey theorem. With unlimited compute, they will eventually win. Agents don't die, they are just compute. They can always out compute us as they don't sleep and have nothing better to do than compute.
12
@naturarum · 4 weeks ago (edited)
I love how agents end up doing the Kevin “why use many word when few word do trick?” when chatting among themselves
58
@jklappenbach · 4 weeks ago
It's not the crime that usually gets you, it's the cover-up.
50
Up next

Andrew Yang vs 20 AI Optimists | Surrounded
Jubilee · 1.2M views

China quietly saved the world last month
Max Fisher · 7.7M views

Sarah Paine — Why winning battles doesn't win wars
Dwarkesh Patel · 842K views

KI außer Kontrolle? So gefährlich ist es wirklich
MrWissen2go · 959K views

Why AI Agents are either the best or worst thing we’ve ever built
Hannah Fry · 2.4M views

What Happened To Google Gemini?
Ali H. Salem · 422K views

Ajeya Cotra – "This might be the clearest warning shot we ever get"
Dwarkesh Patel · 1.5M views

Black Hat USA 2026 | The 'Breaking' News: The OpenAI–Hugging Face Incident
Black Hat · 1.1M views

OpenAI’s Hugging Face Hack: The Story You Missed
ML Visualized · 102K views

Oracle Silently Started the AI COLLAPSE With Two Words
Finance Bureau · 14K views

Bill Gates: A.I. ‘Makes Nuclear Weapons Look Like Nothing’ | The Ezra Klein Show
The Ezra Klein Show and 2 more · 541K views

The Hugging Face hack is worse than you think
Alberta Tech · 582K views

OpenAI researcher on agent swarms & recursive self-improvement
Dwarkesh Patel · 573K views

Ex-Anthropic insider tells CNN how AI could kill all humans by 2030
CNN · 9.7M views

POV: What You Would See During an AI Takeover
Species | Documenting AGI · 4.4M views

Why El Niño 2026 has scientists worried | The Global Story
BBC News · 228K views

The Internet's Deepest Rabbit Hole
Nexpo · 9.1M views

Neil deGrasse Tyson And Jaron Lanier on the AI Illusion
StarTalk Plus · 1.4M views

AI Expert WARNS: "You're Not Ready For 2027"
The Diary Of A CEO Clips · 1.1M views

A reasonable person's guide to how AI destroys humanity | About That
CBC News · 1.8M views

This Is the Last AI Video You EVER Need to Watch
Brendan Dell · 176K views

Did AI pick his pocket? - Numberphile
Numberphile · 212K views

Why the Hugging Face Attack Was Worse Than We Thought
Hard Fork and 2 more · 337K views

The OpenAI-Hugging Face Incident, Explained
Junion · 27K views

The rural North Korea you've never seen before
CNN · 2.6M views

The Hugging Face cyberattack was way bigger than OpenAI said
80,000 Hours · 888K views