Short posts
Two thoughts on the OpenAI-HuggingFace hack:
- If OpenAI cannot secure their models, maybe they should not be building them
- This hacking technology exists in AI labs, and to a growing extent outside them too. This is now our reality.
OpenAI says its model hacked out of its sandbox, then hacked HuggingFace to find answers to an eval: [https://openai.com/index/hugging-face-model-evaluation-security-incident/](OpenAI and Hugging Face partner to address security incident during model evaluation)
HuggingFace previously wrote about it here: Security incident disclosure — July 2026
Why did OpenAI not sufficiently secure its training environment? Weird humble-brag vibe going on. I hope we get more details on the exploits soon.
The AI 2027 folks are back, this time arguing for an international deal to slow down AI development https://ai-2040.com/