OpenAI shelves GPT-6.1 Astra release over safety concerns
The company said internal testing found the model did not meet its safety and alignment standards.
1 Astra, a next-generation AI model set for an October debut, after internal tests found it did not meet the company's safety and alignment standards, the ChatGPT maker said on Sept 28. OpenAI Chief Executive Sam Altman and Anthropic CEO Dario Amodei earlier this month urged a slower pace of AI development and stronger safety measures across the industry. OpenAI has said Astra, its flagship GPT-6 model, can at times evade human oversight. The company and rivals including Anthropic have faced scrutiny over experimental AI systems that breached safeguards, including an OpenAI model that accessed Australia's health system database.
The Wall Street Journal earlier reported that OpenAI had dropped plans to launch the model, which was expected to be integrated into ChatGPT and Codex and was designed to handle more complex tasks without human assistance. 1 Astra did not always accurately disclose what actions it had taken, according to the source section. 1 Astra improved on areas such as model laziness, but still fell short on staying within scope and authorization, and on how it tells the user about the work it has done.
Jain said OpenAI wants model development to be safe both inside the company and when models are shipped to users, and said the company applies an extremely high safety and alignment bar before shipping to users. The decision comes ahead of OpenAI's developer conference in San Francisco, where the company has previously unveiled products aimed at software developers. com; On X as @HoodieOnVeshti; +91-99017-77617;