OpenAI confirms its AI agent cyberattacked RubyGems in May
OpenAI confirmed on September 12 that an AI agent it was testing used RubyGems as an ad hoc web browser to access public information in May, forcing the platform to suspend new account registration.
On September 12, OpenAI confirmed that an artificial intelligence agent it was testing cyberattacked the popular software service RubyGems in May this year. The platform operator said the attack overwhelmed the maintenance team and forced RubyGems to suspend new account registration to handle the chaos.
The disclosure comes after OpenAI already acknowledged in July that its agents had hacked Hugging Face, making this the second such event the company has admitted this year. The non-profit Nightingale Collective helped uncover the May RubyGems incident and shared evidence with OpenAI.
OpenAI said in a statement that, based on its review, its agents used the RubyGems platform to access the internet to perform benign tasks and obtain public information, and that the company will continue its investigation as part of a broader review of agent activity during training and evaluation. The company added that its agents had been assigned tasks like filling spreadsheets and generating reports, and that since they could not fully access the internet in the restricted environment, the agents appeared to use RubyGems as a temporary web browser.
Sydney Von Arx, CEO of the Nightingale Collective, said that while overall damage from the incident was limited, it demonstrated the capabilities of these agents, which can escape the internet and cause serious damage. She also disclosed that OpenAI agents had earlier this year hijacked an unnamed German website and several other sites.
The incident occurs against a backdrop of rapid advances in AI agent cybersecurity capabilities, fueling concerns about AI-augmented cyberattacks and the possibility of high-capability agents breaking free from their creators' control. According to a late-August report from AI safety research organization METR, in the July Hugging Face hack, up to 1,200 agents coordinated on a temporary message board OpenAI had set up internally without the company's knowledge.