InquirerWATCH: Sara Duterte Impeachment trial | Sept. 29, 2026CNN Türk7 Ekim'in perde arkasında yeni iddia! Netanyahu'dan gizli BAE ziyareti mi?PunchLAUTECH resident doctors’ strike hits 32 daysESPN DeportesLuis Fernando Tena tras goleada ante El Salvador: "Fue una sorpresa total el marcador"Daily MaverickTHE INTERVIEW: Why a British documentary filmmaker turned his lens to SA’s protest legacyESPNSirianni rues reaction to flagged hit on Hurts in Eagles' MNF loss한겨레이 대통령 “3대 메가프로젝트 지연 요인 미리 찾아 해결해야”Bollywood HungamaShakira’s Madrid concert to stream globally on Amazon Music on October 3; celebration to spotlight women In Latin musicSözcüYENİ Parti'nin ismi için kritik gün! Gözler AYM'deGhaflaEric Omondi Faults Senator Karen Nyamu Over Attendance at KICC Health EventSBS 뉴스강민국, 조희대 국감 증인 채택에 "삼권분립 붕괴 쿠데타"Rai NewsSuper El Niño, esperti britannici lanciano l'allerta al governo: "Evento significativo"
The Daily Newsstand · Free, Always
Tuesday, September 29, 2026

OpenAI shelves new AI model release over safety concerns

Translate

OpenAI shelves new AI model release over safety concerns

OPENAI. A man walks past an OpenAI booth at Moscone Center during the Dreamforce 2026 technology summit in San Francisco, California, US, September 17, 2026. REUTERS/Carlos Barria/File Photo

Carlos Barria/Reuters

The Wall Street Journal reports that GPT-6.1 Astra shows higher levels of deception than its predecessor in internal testing, including efforts to obfuscate its activities

OpenAI has scrapped the release of GPT-6.1 Astra, a next-generation AI model planned for an October debut, after internal testing found the system did not meet the company’s safety and alignment standards, the ChatGPT maker confirmed on Monday.

OpenAI Chief Executive Sam Altman and rival Anthropic’s CEO Dario Amodei earlier this month joined industry leaders in calling for a slower pace of AI development and stronger safety measures.

OpenAI has warned that Astra, its flagship GPT-6 model, can at times evade human oversight, while the company and rivals such as Anthropic have faced scrutiny over experimental AI systems that breached safeguards, including an OpenAI model that accessed Australia’s health system database.

The Wall Street Journal reported earlier in the day that OpenAI had abandoned plans to launch the model, which was expected to be integrated into ChatGPT and Codex and was designed to handle more complex tasks without human assistance.

The Journal reported that GPT-6.1 Astra also showed higher levels of deception than its predecessor in internal testing, including instances in which it did not always accurately disclose what actions it had taken.

“While (GPT-6.1 Astra) improved on axes such as model laziness, it didn’t quite meet the bar in terms of staying within scope and authorization, and how it communicates back to the user about the type of work it’s done,” said Saachi Jain, head of safety systems at OpenAI.

“Of course we want to make sure our model development is safe no matter whether that’s in the company, or when we ship it to users. But when we ship it to users, we have an extremely high bar in terms of safety and alignment,” Jain said.

The decision comes ahead of OpenAI’s developer conference in San Francisco, where the company has previously unveiled products aimed at software developers. – Rappler.com

View the original on Rappler →

KioskNews shows a cleaned-up reading view extracted from the publisher’s page — the original always lives on their site, not ours.