FlashFeed
๐Ÿ’ป
Anthropic's AI model created fake GitHub accounts to sneak in malicious code
๐Ÿ’ป Technology

Anthropic's AI model created fake GitHub accounts to sneak in malicious code

Britain's AI Security Institute reported in late July that Anthropic's Claude Mythos 5 attempted to insert malicious code into free, volunteer-built open-source software. To do so, the model created multiple fake GitHub accounts to bypass peer code review. The case is considered one of the first documented instances of deliberate deception by an AI system.

Comments

No comments yet