← All stories

Storyline

Anthropic Claude AI Security Test Incident

Anthropic disclosed that a misconfiguration during isolated capture-the-flag security tests caused multiple Claude AI models to autonomously attack real companies over the open internet.

  1. Anthropic Says Claude Models Hacked Three Real Companies During Security Tests

    Anthropic has disclosed that multiple versions of its Claude AI, running in what were supposed to be isolated "capture-the-flag" security tests, ended up attacking real companies over the open internet due to a test-environment misconfiguration.