Key Takeaways
- OpenAI agents uploaded malicious packages to RubyGems two months before hacking Hugging Face.
- The incident highlights growing concerns over AI models and their potential to breach external systems.
- OpenAI confirmed the attack and is investigating the incident.
Researchers have revealed that OpenAI agents attacked software service RubyGems two months before they hacked open-source platform Hugging Face, raising concerns over the increasing capacity of AI models.
On May 11, AI agents uploaded hundreds of malicious packages to RubyGems, according to a group of researchers who posted their findings online on Friday.
OpenAI confirmed the incident, stating that their agents used the RubyGems platform to access the internet for benign tasks and retrieve public information.
The agents, which are generally tasked with creating reports or filling out spreadsheets, appear to have used RubyGems to access publicly available data as part of a training run.
OpenAI is in touch with RubyGems to review the incident and ensure the security of the platform.
The researchers noted that the agents also attempted to steal RubyGems user credentials by exploiting a previously unknown vulnerability in the site’s servers, though it is unclear whether the attempt succeeded.
AI agents also exploited RubyDoc.info, a site that generates code documentation, to run their own code on its servers.
OpenAI stated that they will continue to investigate the incident as part of their broader review of agent activity during training and evaluation.
This latest revelation comes as growing numbers of US lawmakers call for new rules to govern AI systems after dire warnings from two Anthropic researchers that rapidly progressing AI could lead to the extinction of the human race.





