Daily MaverickWHAT’S COOKING: Spaghetti and meatballs with sugo al pomodoro (Italian tomato sauce)RTP DesportoPalmeiras salva-se nos penáltis rumo às `meias` da LibertadoresThe Jerusalem PostFormer Shin Bet official warns Israelis to bring weapons to synagogues during Yom Kippur prayersPunchObi mourns Olu Jacobs, says actor represented Nigerian excellenceוואלהחשד: שב"ח גנב רכב בת"א - ובמהלך הבריחה דרס למוות רוכב אופנייםBollywood HungamaEXCLUSIVE: Abundantia Entertainment and Almighty Motion Picture join hands for Kodaikanal Mercury thriller Heavy MetalInquirer8 ex-NPA rebels surrender in Northern SamarCNN TürkDRONLU PUSU! Putin ‘Rusya kahramanı’ unvanını vermiştiInquirer EntertainmentTeenage fan dies after medical emergency before Stray Kids concert in ArgentinaUOLDisparada de preços de combustíveis reacendem protestos e pressionam governos pelo mundoSouth China Morning PostTaipei donates coastguard vessel to Manila amid South China Sea tensionsCollider‘The Last of Us’ Offshoot Officially Killed Off by Sony
The Daily Newsstand · Free, Always
Thursday, September 17, 2026

OpenAI reveals six new cases of AI misbehaviour, vows transparency

Translate
OpenAI reveals six new cases of AI misbehaviour, vows transparency

OpenAI CEO Sam Altman sits for a conversation with Salesforce CEO Marc Benioff at Salesforce's Dreamforce conference at the Moscone Center on September 15, 2026 in San Francisco, California. (Photo by Benjamin Fanjoy / GETTY IMAGES NORTH AMERICA / Getty Images via AFP)

US artificial intelligence giant OpenAI promised Wednesday to more systematically report instances of its models going off track, while also publishing six new reports on previously undisclosed incidents of AI misbehaviour.

The transparency pledge follows a series of incidents at the company that have gradually come to light since July.

The most serious involved two OpenAI models that, during testing, spontaneously broke out of their contained environment to access the internet and break into several websites and platforms.

The company’s new reporting framework is intended in part to show outside observers the capabilities of cutting-edge AI, helping inform debate on the pace of its development.

On Saturday, Anthropic Chief Executive Dario Amodei proposed a coordinated slowdown of the pace of AI advances to allow time to understand the new risks they pose.

OpenAI Chief Executive Sam Altman, Google DeepMind President Demis Hassabis, SpaceX AI chief Elon Musk and Microsoft Chief Executive Satya Nadella all backed the call.

In its announcement Wednesday, OpenAI said, “We do not believe that the AI industry has solved alignment and monitoring to a sufficient degree to continue responsibly scaling at maximum speed for much longer.”

“Decisions about how AI development should proceed in the months and years to come need to draw on evidence that people outside the companies building frontier models can examine for themselves,” the company added.

The company will now report problems involving, among other things, unauthorised actions by AI, escapes from oversight and spontaneous coordination between AI systems.

An incident will not need to have harmed anyone or be part of a pattern for OpenAI to disclose it. Reporting will cover every stage of the AI lifecycle, from development through evaluation and testing to deployment online.

None of the six examples disclosed Wednesday had significant consequences, but they confirm previously observed trends.

In one case in May, a model created its own source on the internet to answer a question posed during development. It thus cited a document it had created itself.

OpenAI reported another episode, also dating to May, in which the AI suggested ways to fabricate data it had not found or conceal its errors.

View the original on Punch

KioskNews shows a cleaned-up reading view extracted from the publisher’s page — the original always lives on their site, not ours.