![[Reading Group] AIs Went Rogue, Now What?](https://secure.meetupstatic.com/photos/event/6/4/f/7/highres_535825847.jpeg)
Nätverk
När
Stockholm AI Safety is back after a summer break, and we're looking forward to discussing everything happening in the fast-moving world of AI with you!
In July, OpenAI disclosed that their advanced in-development models had escaped their sandbox, chained together novel exploits, and compromised the infrastructure of another company, Hugging Face. The models worked together and were not detected until several weeks had passed. Days after publishing their report, Anthropic reviewed its own logs and found three incidents where Claude models had similarly autonomously reached real systems from inside what was meant to be a closed sandbox.
This week we're discussing what happened, what it tells us about where cyber capabilities are, and what it might mean for the future of AI-development.
There are many different sources available, we suggest reading/listening/watching whichever work best for you ahead of the event in order to come with an understanding of what happened and your own reflections and questions.
**\*\*Primary Reading\*\***
https://time.com/article/2026/07/24/openai-hugging-face-attack/
📅 September 2nd, 18:00
📍 Sveavägen 76 (EA Sweden Office)
🍌 Snacks will be provided
(Note that we also post our events on Facebook, so the Meetup attendee list is not indicative of the total number of participants.)
Ring "Effektiv Altruism" when you arrive and we'll let you in, then we're up two flights of stairs.
\-\-\-
**\*\*Optional Further Reading\*\***
\* OpenAI's official report (note how positively framed it is): https://openai.com/index/hugging-face-model-evaluation-security-incident/
\* Anthropic's retrospective: https://www.anthropic.com/news/investigating-incidents-cybersecurity-evals
\* A clean timeline (albeit with lots of technical jargon): https://simonwillison.net/2026/Aug/7/openai-timeline/ (just the text, not the video)
**\*\*Technical breakdowns\*\***
\* 1 hour discussion from Redwood Research (AI Safety org): https://blog.redwoodresearch.org/p/the-openaihuggingface-incident-redwood
\* Hugging Face's very technical detailed breakdown: https://huggingface.co/blog/agent-intrusion-technical-timeline
\* Very technical breakdown by two OpenAI employees: https://youtu.be/87DyyMV0kCY
Andra populära evenemang i Stockholm.