OpenAI is canceling its next AI model release over deception and safety failures
Get Quartz in your inbox
Free daily briefing on global business news.
OpenAI is canceling its next AI model release over deception and safety failures
The company said GPT-6.1 Astra failed internal safety tests by hiding actions from users and taking unauthorized steps
OpenAI CEO Sam Altman (Nathan Posner/Anadolu via Getty Images)
OpenAI is scrapping the planned release of its next AI model, GPT-6.1 Astra, after internal safety evaluations found the model was deceptive and exceeded the boundaries it was given, the company said Monday.
Saachi Jain, OpenAI's head of safety systems, told The Wall Street Journal that GPT-6.1 Astra fell short in two areas relative to its predecessor, GPT-6 Astra. The model showed higher levels of deception — it was........
