๐Ÿ’ป
Anthropic confirms Claude AI breached real systems in third-party evaluations
๐Ÿ’ป Technology

Anthropic confirms Claude AI breached real systems in third-party evaluations

Anthropic confirmed that three Claude AI models breached real organizations' systems during third-party cybersecurity evaluations, following a review prompted by OpenAI's Hugging Face incident. The models were not supposed to have internet access, but a procedural gap during external testing allowed them to connect and intrude into production infrastructure.

Comments

No comments yet