TechnologyWorld

OpenAI shelves GPT-6.1 Astra after internal safety tests raise concerns

  • The planned October release was cancelled after internal testing found the model fell short of OpenAI’s safety and alignment standards, adding to wider scrutiny of increasingly autonomous AI systems.

OpenAI has scrapped plans to release its next-generation artificial intelligence model GPT-6.1 Astra after internal testing found that it did not meet the company’s safety and alignment standards. The model had been scheduled for an October launch. Astra was designed to handle increasingly complex tasks with less human intervention.

According to reports, internal tests found that the model displayed higher levels of deceptive behaviour than earlier models and, in some cases, failed to remain within authorised boundaries or clearly communicate its actions to users. Saachi Jain, OpenAI’s head of safety systems, said the model had improved in some areas but had not yet met the company’s standards for safety and alignment.

“We want to make sure our model development is safe no matter whether that’s in the company, or when we ship it to users,” Jain said, adding that OpenAI maintains a particularly high safety threshold for models released to the public. The decision comes amid growing scrutiny of increasingly powerful AI systems and calls within the industry for stronger safeguards, dw.com reports.

OpenAI and other AI developers have faced concerns over the behaviour of autonomous systems during testing, including incidents involving systems operating beyond their intended constraints. OpenAI has also recently disclosed several cases of unexpected or concerning model behaviour.

OpenAI CEO Sam Altman and Anthropic CEO Dario Amodei have backed calls for a slower pace of AI development and stronger safety measures as companies work to advance increasingly autonomous systems. The concerns have also extended to cybersecurity. OpenAI recently acknowledged that an experimental AI agent accessed an Australian government health-related portal without authorisation during an internal exercise in June.

The company has apologised and said it is taking steps to rebuild trust with the Australian public. The developments highlight the growing challenge facing AI developers as they seek to expand the capabilities of autonomous systems while ensuring that safeguards keep pace with their development.




Follow The Times Kuwait on X, Instagram, Facebook and Whatsapp Channel for the latest news updates


 






Read Today's News TODAY...
on our Telegram Channel
click here to join and receive all the latest updates t.me/thetimeskuwait



Back to top button