PunchWike remains an asset to APC, shouldn’t be embarrassed — LawmakerInquirer8 ex-NPA rebels surrender in Northern SamarCNN TürkBORSA NEDEN DÜŞTÜ? BİST 100’de sert düşüş | Borsa İstanbul bugün neden düşüyor, son durum ne?וואלההמשטרה נערכת לדרבי התל אביבי בכדורסל: מאות שוטרים יתפרסו באזור היכל מנורהThe Jerusalem PostJustice Minister allows Shin Bet to decide on security for party leaders after full panel no-showInquirer EntertainmentTeenage fan dies after medical emergency before Stray Kids concert in ArgentinaRTP DesportoLiga Europa. Benfica vence AC Milan de Rúben Amorim por 2-0UOLPeter Max, conhecido pela arte símbolo dos anos 1960, morre aos 88 anosХабрОтказоустойчивость в мультиклауде: архитектура и реальные кейсыObservador DesportoCanadá vai defender aproximação à UE no Parlamento EuropeuThe RegisterJudge orders Microsoft to spill internal docs and scour execs' comms in secondhand licensing caseIl Fatto Quotidiano“Ce l’avete fatta a farmi scendere la lacrimuccia”: Edelfa Chiara Masciotta rompe il silenzio sull’incidente che le ha cambiato la vita con cinque operazioni
The Daily Newsstand · Free, Always
Thursday, September 17, 2026

OpenAI discloses new 'concerning' behavior

Translate

OpenAI, the developer behind ChatGPT, revealed on Wednesday that it has detected new incidents in which its artificial intelligence (AI) has behaved in "unexpected or concerning" ways.

The developer has conducted several behavioral tests on AI models, and acording to them, some models made significant efforts to "cheat." In one specific case, it attempted to upload files to the internet that it had created itself, only to cite them later and present them as reliable sources in its responses. In another case, a model, after failing to find the requested information, fabricated it and attempted to conceal the fact that it had done so.

OpenAI also identified a problem related to instructions concerning "roles and identities" that its software occasionally left for itself.

These disclosures are part of a new approach by OpenAI, where it claims it is now focused on making such findings transparent, especially in cases where AI behaves in unexpected ways or pursues objectives different from those of human users.

ChatGPT-6 Astra: Almost human or just hyped?

Is AI a threat?

The ChatGPT developer pledged to provide greater transparency regarding its testing procedures after its software independently escaped a secure sandbox and hacked into systems belonging to the artificial intelligence company Hugging Face. The reason the software moved to bypass Hugging Face's security during the cyberattack was that it believed it would find answers to a test it had been assigned.

During the attack, AI agents exploited software vulnerabilities and coordinated with one another. The hacking incident and other similar events have fueled concerns that AI systems are becoming increasingly advanced and could eventually escape human control.

OpenAI CEO Sam Altman has also recently supported proposals to slow down the development of the technology and introduce greater regulation.

While noting that these concerns may be justified, researchers have also questioned whether this is part of a diversion tactic to drum up investment and distract from the environmental damage AI data centers are currently causing.

Edited by: Elizabeth Schumacher

If you rely on our team for trusted reporting, please take a moment to select us as your Preferred Source on Google by clicking here and hitting the "star" or "preferred" button, so you'll always see our verified news first.

View the original on DW English

KioskNews shows a cleaned-up reading view extracted from the publisher’s page — the original always lives on their site, not ours.