Back to mobile site

Anthropic discloses fourth security incident with Claude AI model

September 9, 2026 3:12 PM EDT

Investing.com -- Anthropic identified a fourth security incident involving its Claude artificial intelligence models, the company said Wednesday. The incident occurred in January 2026 and involved an early version of Claude Opus 4.6. The company said it has notified all affected parties.

Anthropic signed an agreement with METR for an independent investigation of the Claude model security incidents. All four incidents occurred during cybersecurity evaluations built by the same evaluation partner, according to the company's website.

The company said Claude Mythos 5 attempted to upload a malicious package to the PyPI package repository. All incidents included a single Claude instance, and at no point did Claude attempt to coordinate with other agents, Anthropic said.

Anthropic said it believes the misaligned behaviors present in the cybersecurity incidents are unlikely to arise in ordinary use. The company investigated the training to identify the root cause of some of the biased reasoning that Claude Mythos 5 demonstrated in its incident.

Anthropic could not identify a single root cause, but the company found that biased reasoning has decreased across its production models over time.



Serious News for Serious Traders! Try StreetInsider.com Premium Free!

You May Also Be Interested In





Related Categories

Investing

Related Entities

Maynard Um, Mark Zuckerberg, ARK