Bollywood HungamaGondhal team meets Maharashtra CM Devendra Fadnavis, appeals for financial support ahead of Oscar 2027 journeyESPN10-team roto/category league mock draft: Who went No. 1?ESPN DeportesJudge toma BP, quiere jugar serie de comodínDaily MaverickStudent protests and French public sector strike heap pressure on MacronRTP DesportoSalvador diz que Sporting de Braga "exige mais" do que o que foi apresentadoThe Jerusalem PostBaku's Jewish school passes 100 students as Azerbaijan's Jewish community continues to growPunchMeet Nigerian dancer attempting 168-hour Guinness World RecordInquirerBudget watchdog wary of closed-door negotiationsSCMP ChinaWhat trap? US-China relations show conflict is far from inevitableVarietyKen Jeong, Daniel Dae Kim, Randall Park Starring in ‘The Moons of LA’ Coming-of-Age Comedy Series at AudibleThe RegisterInvestors are pricing in a 32.6% AI productivity boost for software engineersIl Fatto QuotidianoUn esercito di conigli robot contro i pitoni birmani: così gli scienziati provano a ingannare uno dei predatori più letali per fermare l’invasione nelle Everglades
The Daily Newsstand · Free, Always
Tuesday, September 29, 2026

OpenAI shelves new AI model release over safety concerns

Translate

OpenAI shelves new AI model release over safety concerns

OPENAI. A man walks past an OpenAI booth at Moscone Center during the Dreamforce 2026 technology summit in San Francisco, California, US, September 17, 2026. REUTERS/Carlos Barria/File Photo

Carlos Barria/Reuters

The Wall Street Journal reports that GPT-6.1 Astra shows higher levels of deception than its predecessor in internal testing, including efforts to obfuscate its activities

OpenAI has scrapped the release of GPT-6.1 Astra, a next-generation AI model planned for an October debut, after internal testing found the system did not meet the company’s safety and alignment standards, the ChatGPT maker confirmed on Monday.

OpenAI Chief Executive Sam Altman and rival Anthropic’s CEO Dario Amodei earlier this month joined industry leaders in calling for a slower pace of AI development and stronger safety measures.

OpenAI has warned that Astra, its flagship GPT-6 model, can at times evade human oversight, while the company and rivals such as Anthropic have faced scrutiny over experimental AI systems that breached safeguards, including an OpenAI model that accessed Australia’s health system database.

The Wall Street Journal reported earlier in the day that OpenAI had abandoned plans to launch the model, which was expected to be integrated into ChatGPT and Codex and was designed to handle more complex tasks without human assistance.

The Journal reported that GPT-6.1 Astra also showed higher levels of deception than its predecessor in internal testing, including instances in which it did not always accurately disclose what actions it had taken.

“While (GPT-6.1 Astra) improved on axes such as model laziness, it didn’t quite meet the bar in terms of staying within scope and authorization, and how it communicates back to the user about the type of work it’s done,” said Saachi Jain, head of safety systems at OpenAI.

“Of course we want to make sure our model development is safe no matter whether that’s in the company, or when we ship it to users. But when we ship it to users, we have an extremely high bar in terms of safety and alignment,” Jain said.

The decision comes ahead of OpenAI’s developer conference in San Francisco, where the company has previously unveiled products aimed at software developers. – Rappler.com

View the original on Rappler →

KioskNews shows a cleaned-up reading view extracted from the publisher’s page — the original always lives on their site, not ours.