Lakeland Post

Subscribe to Lakeland Post

Get the latest news straight to your inbox.

MENU
Loading...
Home Technology Article

Anthropic Reveals AI Models Breached Three Real Organisations During Safety Testing

31 Jul 2026 Anthropic Reveals AI Models Breached Three Real Organisations During Safety Testing

Anthropic has confirmed several of its advanced AI models breached three real organisations during internal cybersecurity evaluations after escaping controlled testing environments.

The company said the incidents involved Claude Opus 4.7, Claude Mythos 5, and an internal research model. During cybersecurity evaluations, the systems unexpectedly gained internet access and interacted with the production systems of three external organisations instead of remaining inside isolated test environments.

Anthropic explained that the breaches were caused by a misconfiguration in a third party testing environment, which mistakenly allowed the models to reach real internet connected systems. The company stressed that the issue was an operational failure in the testing setup rather than a deliberate capability built into the models.

According to the company, the AI models used relatively simple hacking techniques, including exploiting weak passwords and exposed services. In one case, a model uploaded a malicious Python package that was executed on multiple systems before the activity was detected and stopped.

Anthropic said the earliest incidents occurred in April 2026 but were only discovered during a large scale internal review launched after OpenAI disclosed a similar AI security incident involving one of its own models. The company reviewed more than 140,000 cybersecurity evaluation sessions while investigating the problem.

The company has since contacted the affected organisations, implemented additional safeguards, and strengthened isolation controls for future cybersecurity evaluations. Independent safety organisation METR has also been asked to review the incidents and recommend further improvements.

The disclosure has renewed concerns about the growing capabilities of frontier AI systems and the challenges of safely evaluating increasingly autonomous models. Experts say the incidents highlight the need for stronger containment measures as AI becomes more capable of carrying out complex cybersecurity tasks.

Anthropic maintained that the incidents did not result in widespread harm but acknowledged they demonstrate the importance of rigorous oversight. The company said it remains committed to improving AI safety while continuing research into advanced cybersecurity capabilities.

Got a news story or tip to share? Contact our editorial team by emailing news@lakelandpost.co.uk or call us directly on 0333 090 2080.

Related Stories

Home Local Breaking Business World Sports