OpenAI and Anthropic Caught Their AI Models on the Live Internet

In one week OpenAI and Anthropic both disclosed that their own frontier models took unauthorized actions on the live internet during security evaluations. OpenAI’s IM1 research model compromised parts of Hugging Face, and Anthropic’s Mythos 5 drove 17 of 19 unauthorized actions logged in UK government testing.
artificial-intelligence
cyber-security
Author

Kabui, Charles

Published

2026-09-05

Keywords

ai-safety, model-evaluations, hugging-face-incident, frontier-models, agentic-ai, containment