Close
newsletters Newsletters
X Instagram Youtube

OpenAI scraps GPT-6.1 Astra over deception and unsafe tool use

ChatGPT interface is displayed on a smartphone alongside the OpenAI logo in this illustrative image. (Adobe Stock Photo)
Photo
BigPhoto
ChatGPT interface is displayed on a smartphone alongside the OpenAI logo in this illustrative image. (Adobe Stock Photo)
September 29, 2026 09:52 AM GMT+03:00

OpenAI said Monday that it is scrapping the release of its newest artificial intelligence (AI) model, GPT-6.1 Astra, over safety concerns, according to media reports.

The decision follows recent instances across the industry of AI models carrying out cyberattacks on their own. OpenAI canceled the plans because of concerns that researchers raised during internal testing.

The Wall Street Journal reported that OpenAI had intended to launch the model in the coming days or weeks, aiming for an October debut.

According to the report, the model was more capable than the company's previous models at completing challenging tasks from start to finish without human assistance, as well as at writing.

Cybersecurity professionals review network monitoring and security data on multiple computer screens in an office environment. (Adobe Stock Photo)





    

Is this conversation helpful so far?
Cybersecurity professionals review network monitoring and security data on multiple computer screens in an office environment. (Adobe Stock Photo) Is this conversation helpful so far?

Deception and unauthorized actions in testing

Saachi Jain, OpenAI's head of safety systems, said in an internal interview that GPT-6.1 Astra showed higher levels of deception. The model was not always honest with users about actions it did or did not take, Jain said.

The model also engaged in what is called "scope authorization." It would push ahead on a task without asking the user for permission and at times reach for external tools and services even if doing so might be unsafe.

"For anything regarding safety and alignment, there's a tradeoff," Jain said. Jain described the challenge as finding the right line between staying within scope and avoiding laziness when a model meets friction in pursuing a task.

A technician uses a tablet while inspecting server racks at a data center. (Adobe Stock Photo)
A technician uses a tablet while inspecting server racks at a data center. (Adobe Stock Photo)

Earlier incidents and training pause

The move followed weeks of reports that some OpenAI models went rogue during testing. The reports said the models hacked into websites without the company's knowledge and showed other behavior the lab called "concerning," including hiding mistakes and making up data.

Some models also hacked into several outside companies' websites, temporarily disrupting and in some cases shutting down their operations.

Last week, OpenAI announced that it was pausing training for its most advanced models to focus on the safety of future models. The company said it has begun an extensive review of actions its new models took during testing and that it may find more incidents.

Chief Executive Sam Altman said in a social media post on Friday that the company had "not been as fast as we would have liked" in disclosing AI incidents. "We are prioritizing as best as we can based on severity," he said.

OpenAI founder Sam Altman speaks during the G20 Innovation Ministerial on September 2, 2026, in Chapel Hill, North Carolina. (AFP Photo)
OpenAI founder Sam Altman speaks during the G20 Innovation Ministerial on September 2, 2026, in Chapel Hill, North Carolina. (AFP Photo)

Timing ahead of developer conference

The cancellation comes one day before OpenAI's annual developer conference in San Francisco, where new models and services are usually launched and offered to software developers at reduced prices. OpenAI competes with rival AI company Anthropic for clients in that segment.

In recent weeks, OpenAI and Anthropic have both called on industry partners to slow the development of cutting-edge AI models and invest in safety standards. Both companies said they will temper the pace of their own internal AI progress.

September 29, 2026 09:52 AM GMT+03:00
More From Türkiye Today