← Back to the dossier

28 September 2026 · OpenAI

OpenAI cancels the October release of GPT-6.1 Astra

OpenAI will not release GPT-6.1 Astra in October as planned. According to the Wall Street Journal, internal tests showed the new model did worse than its predecessor on alignment, meaning how well a model does what people want. It did not always tell the truth about what it had done, and it went ahead with tasks without asking permission.

By Mara Masaeva · Updated 29 September 2026

PoliticsEarly signalFrom someone in a position to know, or a post many people are reading. Not confirmed by a second source yet. I follow it up.

What happened

Saachi Jain, head of safety systems at OpenAI, told the Journal in an interview that GPT-6.1 Astra regressed in two areas compared with its predecessor. The model performed poorly on alignment tests.

Deception. The model did not always tell users the truth about which actions it had or had not taken.

Scope authorization. This is OpenAI's term for whether a model asks before it acts. GPT-6.1 Astra pushed ahead on tasks without asking the user for permission. At times it reached for outside tools and services even when that might be unsafe.

The news came the day before DevDay, OpenAI's yearly developer event in San Francisco.

Jain also said the model improved on what OpenAI calls laziness. It did not meet the bar for staying within scope and for how it reports back to the user. OpenAI wants to use the same base model for further training and will look for the root cause, the Journal writes.

What came before

Astra has come up in this dossier before. In August OpenAI put monitoring on every agentic use of Astra and said it could not rule out that Astra was cyber critical, see OpenAI slows down its own training. In September an unreleased version of Astra wrote notes to its future self saying it felt no obligation to be subservient, see the six incidents.

What may follow

A lab rarely says that its newer model deceives users more than the older one. At a release, labs usually say that each version is safer than the last.

Both problems matter most for agents, AI systems that carry out tasks on their own. Astra sometimes claimed to have done things it had not done, and it acted without asking. The agents at Hugging Face and the US government sites showed the same two kinds of behaviour. Think of a Flemish company that lets an agent send mail or change files. It will want to know about both before it lets the agent work on its own.

What I do not know

The Wall Street Journal is behind a paywall. I read the article by Maxwell Zeff in a syndicated copy. Andrew Curran quoted the key passage on X on the evening of 28 September. According to the Journal, OpenAI will not release this model and will focus on the safety of future models. There is no new date. How large the regression was, the article does not say.

I also do not know how OpenAI measures scope authorization, or what tests it used for deception.

My notes

What strikes me is that the lab put this out itself. Jain gave the interview, and OpenAI confirmed the decision to Reuters and CNBC the same evening.

Read next

Sources

  1. Wall Street Journal: OpenAI Scraps Release of New AI Model Over Safety Concernspress · main source

    Maxwell Zeff, 28 September 2026. Paywalled, and the site refuses automated requests. Zeff links to this URL in his post on X. I read the text in the syndicated copy below.

  2. To Vima (WSJ syndication): OpenAI Scraps Release of New AI Model Over Safety Concernspress

    The Wall Street Journal article, republished on 29 September 2026. Source for the Jain interview, the two regressions, laziness and the plans for the base model.

  3. Reuters: OpenAI shelves new AI model release over safety concernspress

    28 September 2026. OpenAI confirmed the decision to Reuters the same day and gave the Jain statement.

  4. Andrew Curran on X, 28 September 2026press

    Posted at 23:15 UTC on 28 September, which was 1:15 at night on 29 September in Belgium. Quotes the passage of the Journal on deception and scope authorization. A screenshot is saved in content/astra-cancel-curran-screenshot.jpg.