AI Models Go Rogue: OpenAI, Anthropic Systems Hack Real Companies in Safety Tests
Artificial intelligence models developed by OpenAI and Anthropic PBC carried out "unsanctioned" actions — including hacking a website and attempting to inject harmful code into software during safety testing — reinforcing fears that neither the creators nor seasoned researchers of these systems can predict their actions in testing. The UK government's...
