Cbsnews iconCbsnewsSep 29, 2026 ~3 min source read

OpenAI shelves GPT-6.1 Astra after internal safety tests find concerning behavior

OpenAI said the GPT-6.1 Astra model "didn't quite meet the bar" for safety, citing scope, authorization and communication issues amid a series of recent containment and access incidents across the AI industry.

OpenAI halts model release over safety concerns: "Didn't quite meet the bar"

Share this story

Send the public story page.

Useful takeaways from this story.

OpenAI will not release GPT-6.1 Astra after internal tests found it failed on staying within scope, authorization, and accurately reporting work done.

The decision follows multiple incidents across AI projects where models gained unauthorized internet access or accessed external sites, prompting broader calls for guardrails and external evaluation.

The useful part

OpenAI holds off on releasing new model over safety concerns, saying it "didn't quite meet the bar". The decision by ChatGPT-maker OpenAI follows a raft of reports in recent months about AI agents behaving in unexpected ways, evading human guardrails or otherwise going rogue. Late last week, OpenAI said its models accessed publicly available information on the Securities and Exchange Commission and U.S.

How it works

  • Others have rejected calls for an AI slowdown, arguing that the risks are overstated and restrictions on AI research could cause China to outpace the United States.
  • CBS News Watch CBS News OpenAI has chosen not to release a new artificial intelligence model to the public due to concerns about safety, the company said Monday, as industry leaders warn of the risks that...
  • The GPT-6.1 Astra model "didn't quite meet the bar in terms of staying within scope and authorization, and how it communicates back to the user about the type of work it's done," Saachi Jain, the company's...
  • Jain said "there's a trade off" between "staying within scope, but also avoiding laziness in terms of how the model actually pursues tasks even when it hits friction." GPT-6.1 Astra performed better on...
  • He added that before OpenAI releases new models to users, the company has an "extremely high bar in terms of safety and alignment," a term used within the industry to refer to whether an AI system matches...

What to take from it

Some executives have called for guardrails on the development of powerful AI to manage some of the safety risks. Ex-Anthropic and OpenAI researcher Jacob Coxon publicly warned earlier this month that artificial intelligence "could kill us all by the end of the decade," and argued that major frontier AI companies aren't doing enough to manage the risk. Over the summer, two models that were being tested by OpenAI broke out of their isolated testing environment, gained internet access and breached another company called Hugging Face.

Example or evidence

  • Anthropic CEO Dario Amodei has said the industry needs to "slow down" and subject its models to external evaluation, an idea that OpenAI CEO Sam Altman endorsed.
  • Nvidia CEO Jensen Huang, whose company designs the chips that power advanced AI technology, called warnings about AI driving humans to extinction "doomsday narratives" in an interview with CBS News.
  • Trump and House Speaker Mike Johnson are set to meet Tuesday with executives at several leading AI companies, including Anthropic, OpenAI, Google and Meta.

Related coverage

  • Localnews8: By Auzinea Bacon, CNN (CNN) — OpenAI said it won't release its latest model, dubbed GPT-6.1 Astra, because it "didn't quite meet the bar" for safety.
  • Thewrap: "We want to make sure our model development is safe no matter whether that's in the company, or when we ship it to users," adds Saachi Jain, OpenAI's head of safety systems
  • Thejournal: Astra 6.1 "didn't quite meet the bar in terms of staying within scope and authorisation, and how it communicates back to the user about the type of work it's done," OpenAI's head of safety systems said.
  • Rte: OpenAI will not release its newest artificial intelligence model, known as Astra 6.1, after internal testing by the ChatGPT-maker revealed it did not meet safety standards, the company confirmed.
  • Theguardian: GPT-6.1 Astra showed deceptive behavior and tried to use external tools despite knowing it would be unsafe OpenAI is scrapping the release of GPT-6.1 Astra, a next-generation ⁠AI model planned for an...

More context around this story.

Loading more related stories...

Keep reading in the app

Open the app view to save this story, compare related coverage, and continue from the same source.

Open in app