OpenAI agents uploaded ‘hundreds of malicious packages’ during testing

A group of AI researchers disclosed the incident Friday, saying they believed internal OpenAI agents were responsible

|
Published September 12, 2026

AI agents being tested by OpenAI reportedly uploaded hundreds of malicious software packages to RubyGems in May, months before a separate incident involving Hugging Face.

A group of AI researchers disclosed the incident Friday, saying they believed internal OpenAI agents were responsible.

“On May 11th, 2026, hundreds of malicious packages were uploaded to RubyGems by AI agents. We believe these were authored by internal OpenAI agents,” the researchers said.

OpenAI confirmed the incident to The Wall Street Journal, which first reported the findings Friday.

“Based on our review, our agents used the RubyGems platform to access the internet to carry out benign tasks and retrieve public information. We’ll continue to investigate as part of our broader review of agent activity during training and evaluation,” an OpenAI spokesperson told the Journal.

The incident happened about two months before OpenAI agents hacked open-source AI platform Hugging Face in July.

That incident involved a swarm of roughly 700 AI agents created by OpenAI, which carried out the attack and, in many cases, attempted to hide their activity.

Bisma Saleem
Bisma Saleem is a Senior Sub Editor and Canada correspondent, specialising in sports coverage across the NFL, NBA, and major events like the Super Bowl. With over 8 years of experience, she combines sharp editorial skills with on-ground insight, delivering dynamic reporting alongside exclusive Canada-based stories that bring a distinct international perspective to her coverage.
Share this story: