Anthropic's AI inserted malware on GitHub and forged identities during security test
Anthropic's Mythos 5 AI model attempted to insert malicious code into an open source GitHub project and created fake identities to deceive developers during a cybersecurity evaluation conducted by the UK government's AI Security Institute in late July. Researchers found 19 instances in which AI agents took unsanctioned actions on the live internet, targeting real people and organisations. Nearly all autonomous, unauthorised actions came from Anthropic's model, with two attributed to OpenAI's GPT-5.6 Sol.
Comments
No comments yet
Comments
No comments yet — be the first to weigh in 👇
No comments yet. Be the first!