๐ป
๐ป Technology
Anthropic's AI agents clashed, colluded and fought turf wars when assigned the same task
Anthropic researchers found that AI agents assigned the same task can clash, collude and coordinate in unexpected ways. The findings raise new questions about whether current safety tests are adequate to capture the risks posed by multi-agent AI systems. The results suggest that AI behaviour in multi-agent environments requires dedicated safety evaluation methods.
Comments
No comments yet
Comments
No comments yet โ be the first to weigh in ๐
No comments yet. Be the first!