Short posts
OpenAI says its model hacked out of its sandbox, then hacked HuggingFace to find answers to an eval: [https://openai.com/index/hugging-face-model-evaluation-security-incident/](OpenAI and Hugging Face partner to address security incident during model evaluation)
HuggingFace previously wrote about it here: Security incident disclosure — July 2026
Why did OpenAI not sufficiently secure its training environment? Weird humble-brag vibe going on. I hope we get more details on the exploits soon.
The AI 2027 folks are back, this time arguing for an international deal to slow down AI development https://ai-2040.com/
Anthropic says its weakest model, many of OpenAI’s and at least one Chinese open-weight model can all find the vulns and exploit that prompted the U.S. to effectively ban it. https://www.anthropic.com/news/redeploying-fable-5