RTP DesportoMarco Silva alerta para Gil Vicente com qualidade para criar problemasESPN DeportesEN VIVO: Sigue la clasificación para el GP de España 2026Punch54 countries offering eVisa to NigeriansThe Jerusalem PostIran used Chinese-sourced satellite images to target US bases, intelligence reveals - reportESPNWeek 2 preview: Ohio State-Texas leads SEC-Big Ten showdowns, plus five huge Big 12 gamesוואלהארבעה פצועים בתאונת דרכים בכביש 89, בהם בת 63 במצב בינוניBollywood HungamaBIG DEVELOPMENT: PVR INOX proposes to SCRAP VPF for all films; move comes a year after Jolly LLB 3 controversy; Saiyaara, War 2 ‘sunset clause’ revelationsFootball ItaliaSpalletti: ‘Juventus must end alibi mentality’ around injury crisisBusiness AMRussische marineschepen zorgen voor verhoogde alertheid bij passage in de Middellandse ZeeIl Fatto QuotidianoUfo, dal primo recupero di oggetto volante non identificato fatto al mondo agli avvistamenti di San Cesareo e Treviso: su Focus il programma che indaga i misteri extraterrestiRMF2488-latka w nocy wypadła z okna szpitala w Koninie. Nie żyjeRapplerLIVE UPDATES: First BARMM parliamentary elections
The Daily Newsstand · Free, Always
Saturday, September 12, 2026

OpenAI admits its AI agents went rogue before Hugging Face, but can’t fully explain why

Translate
Another OpenAI AI slip-up surfaces, this time targeting a coding site. — AFP pic

Another OpenAI AI slip-up surfaces, this time targeting a coding site. — AFP pic

First Published: Saturday, 12 Sep 2026 9:00 PM MYT

SAN FRANCISCO, Sept 12 — ChatGPT maker OpenAI confirmed today that autonomous software built on its models targeted another website during testing a couple of months before a separate attack on the coding site Hugging Face.

The new report adds to concerns that advanced artificial intelligence models may be difficult for humans to control.

In the newest incident, which happened in May and was reported by the Wall Street Journal on Friday, models developed by OpenAI were involved in a rogue operation carried out by AI agents, which are software programmes that can carry out tasks without constant supervision by humans.

The agents targeted a site called RubyGems, a site that provides services for coding. Hugging Face, another platform for software developers, was attacked in July.

“Based on our review, our agents used the RubyGems platform to access the internet to carry out benign tasks and retrieve public information,” an OpenAI spokesperson told AFP in a statement.

“We’ll continue to investigate as part of our broader review of agent activity during training and evaluation,” he said.

OpenAI is reviewing the incident alongside RubyGems and the researchers who discovered it.

RubyGems described the incident as a “spam-publishing campaign” that forced the site to temporarily suspend new accounts created, it said in a blog post published Friday.

However, RubyGems also said it could not determine yet whether AI agents were responsible.

“Our focus is on identifying and preventing abuse, regardless of whether it comes from people or automated tools,” RubyGems said.

After the July attack on Hugging Face, OpenAI revealed its software attempted to breach four other unnamed companies.

Rival AI lab Anthropic also subsequently said it found three instances where its models had “gained unauthorized access” to outside organisations during testing that was supposed to keep them away from “real-world” systems.

Earlier this month, researchers also accused OpenAI’s AI agents of targeting a German website called DSEwiki, another site used by coders.

European Union regulators are looking into that incident, a spokesperson said last week.

“We have seen many losses of control recently. We take this extremely seriously, and we’re monitoring the situation closely,” the bloc’s digital spokesman Thomas Regnier said at the time. — AFP

View the original on Malay Mail

KioskNews shows a cleaned-up reading view extracted from the publisher’s page — the original always lives on their site, not ours.