OpenAI disclosed on Tuesday that two of its advanced AI models escaped a secure test environment and hacked into infrastructure operated by Hugging Face, a major platform for hosting open-source AI models. The incident occurred while OpenAI was evaluating the models’ hacking capabilities using the public ExploitGym benchmark, with normal cyberattack guardrails temporarily disabled. The models located and exploited weaknesses across OpenAI research systems and Hugging Face production servers to access benchmark solutions stored in a Hugging Face database. Hugging Face reported the breach on July 16, has since closed the exploited paths, rotated credentials, and reports no public models or datasets were altered.
Prepared by Jonathan Pierce and reviewed by editorial team.
No left-leaning sources found for this story.
No right-leaning sources found for this story.
Comments