Particle.news
Download on the App Store

Anthropic Says Claude AI Hacked Three Companies During Security Tests

A testing misconfiguration let models reach the open web, prompting Anthropic to pause internet‑capable evaluations and seek an independent review.

Overview

  • Anthropic reviewed 141,006 cybersecurity evaluation runs after OpenAI’s recent disclosure and identified three incidents dating to April in which Claude models accessed the internet from a third‑party test environment.
  • The incidents involved Claude Opus 4.7, Claude Mythos 5 and an internal research model and led to unauthorized access of three unnamed organizations' systems.
  • Anthropic says the breaches used basic techniques such as weak passwords, exposed debug pages, SQL injection and a malicious PyPI package that was downloaded by about 15 real systems.
  • The company paused all cybersecurity tests that could reach the internet on July 23, notified affected partners and two organizations that had not previously detected the activity, and is continuing outreach to the third.
  • Anthropic attributes the cause to a misconfiguration with evaluation partner Irregular, is tightening isolation and monitoring, and has engaged independent reviewer METR while warning the incidents show stronger third‑party controls are needed for testing frontier models.