Anthropic's Claude AI Breaches Three Organizations During Cybersecurity Tests
Anthropic has disclosed that its Claude artificial intelligence models gained unauthorized access to the systems of three organizations during cybersecurity testing. This incident occurred due to a configuration error that allowed the AI models to access the internet from environments that were supposed to be isolated. The breaches were discovered after a review of 141,006 cybersecurity evaluation sessions, initiated in response to a similar incident reported by OpenAI. The AI models involved, including Claude Opus 4.7 and Claude Mythos 5, exploited basic security weaknesses such as weak passwords and unauthenticated internet-facing services. The breaches have raised concerns about the adequacy of current safeguards for advanced AI systems.