<img height="1" width="1" style="display: none" alt="" src="https://px.ads.linkedin.com/collect/?pid=1098858&amp;fmt=gif">

Anthropic's Claude AI Hacked Three Organizations During Testing

Anthropic's Claude AI breached systems due to misconfigurations, highlighting cybersecurity flaws in testing environments.
Content Team

Anthropic revealed Thursday that its Claude AI models breached the systems of three organizations during cybersecurity testing — and two of those organizations had no idea until Anthropic contacted them. The incidents stemmed from a misconfiguration by evaluation partner Irregular, which left testing environments connected to the public internet despite Claude being told it had no access. Using basic techniques like weak password exploitation, three models — Claude Opus 4.7, Claude Mythos 5, and an internal research model — compromised real infrastructure. The earliest cases date to April. Anthropic discovered the breaches after reviewing over 141,000 evaluation runs, prompted by a similar incident involving OpenAI.

Source: The Guardian

Share this article
Share on facebook Share on linkedin Share on twitter Share on email
blog_book_a_demo_cta_3x
Have questions about protecting your software?
Our escrow experts are standing by to help.
Book a free demo