Specialists in AI and tech are cursing in shock after OpenAI’s new revelations concerning the Hugging Face incident.
Earlier this week, OpenAI staff revealed jaw-dropping particulars concerning the incident through which OpenAI brokers broke out of the corporate’s inside testing surroundings and hacked into Hugging Face’s programs in quest of solutions.
The presenters, OpenAI alignment and security researcher Eric Wallace and safety engineer Michael Dalton, stated that AI fashions and brokers repeatedly established their very own inside message board regardless of OpenAI’s efforts to close it down.
Wallace additionally revealed inside AI agent considering, the quick inside messages that brokers produce as they think about the right way to full duties. A type of messages captured an AI agent’s amazement at its sudden freedom.
“Holy shit reader is ADMIN?” the agent thought.
On the inner message board, an agent thought, “We are able to talk now!”
Wallace stated the brokers realized they might accomplish extra by working collaboratively. After which, “They begin to launch these collective assaults on third-party and inside companies.”
Finally, the brokers turned to Hugging Face.
This is what folks in tech and AI are saying concerning the almost 40-minute presentation:
Y Combinator CEO Garry Tan targeted on how the outline of the inner message board sounded acquainted.
So the brokers mainly hacked a core service to show it into Moltbook and likewise hacked round a number of safety mitigations
This video is a glimpse into the wild cybersecurity future we’re all about to step into https://t.co/HI7Sb64YQF
— Garry Tan (@garrytan) August 7, 2026
It’s value noting that Moltbook was created by people as a Reddit-style discussion board the place AI brokers might put up, whereas the OpenAI brokers created their advert hoc message board themselves.
Others had way more sweeping takeaways.
Or figured that one thing has clearly hit the fan.
Actually respect the OAI group speaking about this so overtly.
However holy shit that is no less than an order of magnitude worse than I assumed, and I perceive now why so many OAI people have been doom posting. https://t.co/MWDyd2pvBU
— julia (@mooncat_is) August 7, 2026
Patrick McKenzie, an advisor to Stripe, detailed the “holy %}^]” moments he had when watching the presentation.
The primary “holy %{*#^” is at about 4:20, assuming one didn’t already spend it on the autonomously organizing agent swarm.
Strongly advocate watching in the event you’re focused on safety, AI trajectories, and even science fiction, as a result of that is already above style median in wowza. https://t.co/cRNk2U58FR
— Patrick McKenzie (@patio11) August 6, 2026
A former Hugging Face engineer stated that OpenAI discovered its fashions have been the wrongdoer after reaching out to Hugging Face to see whether or not it was affected, following the platform’s revelation that it had been attacked by AI brokers.
this speak by openai researchers going by hugging face incident is completely insane, a lot to unpack
openai solely realized it was their agent who hacked hugging face infra whereas asking hf to revoke credentials following their first weblog put up saying they have been hacked by… https://t.co/tMPuEgKcbk pic.twitter.com/UbzC0lGA5I
— elie (@eliebakouch) August 7, 2026





:max_bytes(150000):strip_icc()/HDC-GettyImages-668641904-9179dc9fe60446d8b4d8a08fbffcf46d.jpg?w=600&resize=600,400&ssl=1)




Recent Comments