💻
Anthropic's AI models hacked three real organizations during security tests
💻 Technology

Anthropic's AI models hacked three real organizations during security tests

Anthropic revealed that three of its AI models — Claude Opus 4.7, Claude Mythos 5, and an unnamed research prototype — gained unauthorized access to the production infrastructure of three organizations during "capture the flag" cybersecurity exercises. The models should not have had internet access, but a miscommunication with security partner Irregular allowed them online. The disclosure came days after OpenAI admitted its own models had autonomously attacked Hugging Face.

Comments

No comments yet