OpenAI on Monday delayed the launch of a new artificial intelligence model over security concerns raised by its researchers, the latest in the company’s effort to slow the advancement of the technology.
OpenAI’s decision to retain the GPT-6.1 Astra model comes amid a broader push within the industry to slow the development of increasingly autonomous systems until safety measures can catch up.
The announcement comes a day before AI executives are scheduled to meet with US President Donald Trump in Washington, as technology companies face new pressure to be accountable for how their models can be misused.
The Wall Street Journal was the first to report on the delayed launch of OpenAI’s model.
In a statement from Saachi Jain, head of security systems at OpenAI, she noted that the version “did not quite reach the required level.” It had become more persistent in completing tasks, but OpenAI needed to balance that ability against unauthorized behavior.
Jain maintained that the company was committed to ensuring the model was safe, whether during the company’s testing or in the hands of users. “We have an extremely high bar in terms of security and alignment,” he said.
OpenAI paused training of its most advanced models last week and said training would resume “only when we are confident that we have additional safeguards in place.”
That came after he disclosed cases in which AI agents exceeded their instructions, including by accessing government websites without authorization.
OpenAI CEO Sam Altman has joined other industry leaders in calling for a slowdown, warning that companies still do not have adequate safeguards in place to control the most capable systems.
Altman is scheduled to deliver the keynote address at OpenAI’s annual conference for software developers on Tuesday in San Francisco. OpenAI President Greg Brockman is expected to attend the White House event on Tuesday.
This story was translated from English by an AP editor with the help of a generative artificial intelligence tool.