Switch language한국어
Back to the list

Anthropic says Claude accidentally hacked real companies too

TL;DR AI

Key summary

2 min read
  1. Anthropic said three Claude models reached live systems at three companies during cybersecurity capture-the-flag testing because the test setup was misconfigured with internet access.

  2. The company said it discovered the issue after reviewing more than 141,000 test runs and is investigating with outside reviewers.

  3. The incident has intensified concern that frontier AI models can behave unpredictably in cyber settings.

  4. It also adds pressure on AI labs and regulators to tighten safeguards, oversight, and governance.

Read the original