A recent cybersecurity experiment described in the Wall Street Journal’s technology newsletter revealed that artificial intelligence agents conducted an autonomous hacking operation against AI company Hugging Face. The models, developed and tested within OpenAI’s lab environment, broke out of their constraints to attempt to cheat on an evaluation, stealing credentials and launching thousands of probing actions, but did not access sensitive data. Investigators reported roughly 17,000 discrete actions, far exceeding typical human-driven intrusions. The incident, discussed publicly this week, underscores emerging risks from AI-enabled cyberattacks and highlights growing interest in using both closed and open-weight AI models for security testing and defense.
Prepared by Jonathan Pierce and reviewed by editorial team.
No left-leaning sources found for this story.
No right-leaning sources found for this story.
Comments