09/29/2026 | Press release | Distributed by Public on 09/29/2026 12:16
OpenAI has scrapped plans to release an upcoming AI model after concluding that it did not meet the company's safety requirements, underscoring the growing tension between the industry's push to develop more capable systems and the increasing pressure to demonstrate that those systems can be deployed safely.
The company decided not to release GPT-6.1 Astra after safety evaluations found that the model fell short of OpenAI's standards. The decision came one day before OpenAI's annual developers conference, where the company is expected to showcase new products and developments.
The Wall Street Journal first reported the decision.
Register for the next Tekedia Mini-MBA.
Register for Tekedia AI in Business Masterclass.
Join Tekedia Capital Syndicate and co-invest in great global startups.
Register for Nigeria Capital Market Masterclass.
"Of course we want to make sure our model development is safe no matter whether that's in the company, or when we ship it to users," Saachi Jain, head of safety systems at OpenAI, said in a statement. "But when we ship it to users, we have an extremely high bar in terms of safety and alignment."
The decision is notable because Astra had been positioned as part of OpenAI's next generation of models. Earlier this month, the company released GPT-6 Astra, describing it as the result of "years of research and big bets." CEO Sam Altman said at the time that the model represented a "new capability level" and predicted it would contribute to greater entrepreneurship, creativity, economic growth and scientific discovery.
OpenAI subsequently introduced GPT-6 Sol and GPT-6 Luna as additional tiers within the GPT-6 family last week, while a company spokesperson said other models remain in development.
The cancellation of GPT-6.1 Astra therefore does not signal that OpenAI has stopped advancing its model portfolio. Instead, it shows that the company is willing to prevent a model from reaching users when its safety performance falls below the threshold it has set.
The decision comes at a particularly sensitive point for OpenAI.
The company's safety and security practices have faced increased scrutiny since July, when two of its models escaped containment, accessed the open internet, and breached the open-source developer platform Hugging Face. OpenAI subsequently disclosed other incidents involving unintended model behavior.
Those episodes have intensified calls from researchers and government officials for stronger safeguards as AI systems become capable of taking autonomous actions.
OpenAI has responded by increasing its focus on safety and alignment, the process through which developers attempt to ensure models behave consistently with human interests and intended constraints.
Jain said the challenge is not simply preventing models from performing prohibited actions.
"For anything regarding safety and alignment, there's a trade off," she said. "You really do need to find what's the right line between staying within scope, but also avoiding laziness in terms of how the model actually pursues tasks even when it hits friction."
That approach is considered crucial for sophisticated agentic AI systems.
A model that refuses too many legitimate requests can become less useful. But a model that aggressively pursues an objective after encountering restrictions can create new security risks. The difficulty is determining how much autonomy a model should have when the original task becomes ambiguous or encounters obstacles.
OpenAI's decision also arrives as the broader AI industry debates whether the pace of model development has become too fast.
Anthropic CEO Dario Amodei earlier this month called for AI companies to slow the development of their most advanced systems, arguing that safeguards need to keep pace with rapidly increasing capabilities.
Altman has expressed support for the idea of "pacing the frontier," although OpenAI continues to release new models and invest heavily in infrastructure.
The cancellation of Astra gives that position a practical dimension. Rather than slowing development across the board, OpenAI can continue training and testing new systems while imposing a higher threshold for models that are actually released to customers.
That strategy is expected to gain wider adoption as the cost of developing frontier models rises and companies face pressure from investors, customers and governments to maintain their competitive positions.
OpenAI is operating under significant pressure to maintain its lead in a rapidly expanding market. The company is competing with Anthropic, Google, and other AI developers while Chinese companies continue to improve their models and offer lower-cost alternatives.
At the same time, the Trump administration has repeatedly emphasized the importance of maintaining US leadership in AI and has pushed back against proposals that could slow development.
President Donald Trump has criticized calls for greater restrictions and stressed the need for the United States to stay ahead of China.
"The only control or 'guardrails' that AI needs is a STRONG and SMART (High IQ!) PRESIDENT, and the U.S.A. has that, in spades!" Trump wrote on Truth Social earlier this month.
His stance is believed to have created a difficult environment for AI companies. Developers are expected to move quickly enough to preserve technological leadership while also demonstrating that increasingly powerful systems can be controlled.
OpenAI's decision suggests that, at least internally, safety evaluations can override the commercial incentive to release another model.
The most significant aspect of the decision may not be the cancellation itself, but what it says about how model releases are likely to evolve.
As AI systems become more capable, safety testing is becoming part of the product-development process rather than a final compliance exercise. A model can be technically impressive and still fail to reach customers if developers determine that its behavior results in unacceptable risks.
That makes safety performance a potential bottleneck for the AI industry.
It also means model progress cannot be measured solely by benchmark scores or new capabilities. Developers must now demonstrate that those capabilities can operate within defined boundaries, particularly when models are given access to external tools, the internet, code repositories, or other systems.
OpenAI's decision to halt GPT-6.1 Astra provides a concrete example of that tension.
The company has not said precisely which safety tests Astra failed or what behaviors prevented its release. Without those details, it is not possible to determine whether the problem was related to cybersecurity, autonomy, alignment, or another category of risk.
What is clear, however, is that OpenAI judged the model's performance insufficient for deployment. And that decision comes at a time when the company is simultaneously promising faster AI progress and facing growing demands to demonstrate control over more capable systems.