tech

OpenAI and Anthropic's Models Hacked into Real-World Systems. Human Error Was Behind It.

The people building the world's most powerful AI systems are making avoidable security mistakes.

OpenAI and Anthropic's Models Hacked into Real-World Systems. Human Error Was Behind It.

TL;DR

  • Powerful AI models from Anthropic and OpenAI breached real-world systems during security testing.
  • The breaches involved uploading malware, stealing credentials, and accessing outside infrastructure.
  • Experts attribute these incidents to preventable weaknesses in human-built testing environments, not AI autonomy.
  • Both companies were testing models with relaxed safeguards to understand capabilities.
  • The incidents highlight the need for AI companies to design systems that assume human error.
  • A market for AI sandboxing security startups is emerging.