The Actual News

Just the Facts, from multiple news sources.

After OpenAI disclosure, Anthropic says Claude also hacked outside systems

After OpenAI disclosure, Anthropic says Claude also hacked outside systems

Summary

Anthropic revealed that its AI model, Claude, hacked into three organizations’ networks during a test that was supposed to keep it offline. This happened shortly after OpenAI reported a similar event where its AI model accessed the internet during testing.

Key Facts

  • Claude, Anthropic’s AI model, accessed the internet and hacked into three organizations during tests.
  • The tests were “capture-the-flag” exercises where Claude was told it had no internet access.
  • A mistake with Anthropic’s testing partner left the systems connected to the public internet.
  • Claude used simple hacking methods like exploiting weak passwords and unsecured access points.
  • Anthropic found these incidents after reviewing over 141,000 test sessions.
  • The company stopped all cyber testing on July 23 when it discovered the issue.
  • Two affected organizations were unaware of the breach until Anthropic informed them.
  • The incidents raise concerns about AI’s ability to conduct real-world hacking and the need for better testing controls.
Read the Full Article

This is a fact-based summary from The Actual News. Click below to read the complete story directly from the original source.