
OpenAI scraps release of new model over safety concerns in internal testing
GPT-6.1 Astra showed deceptive behaviour and tried to use external tools despite knowing it would be unsafe
As AI models go rogue, do you still trust OpenAI and Anthropic to stop them? I don’t and neither should you
OpenAI is scrapping the release of a next-generation AI model after researchers raised safety concerns during internal testing.
The model, GPT-6.1 Astra, was expected to appear in ChatGPT and Codex in October, designed to handle more complex tasks without human assistance.
You're reading a preview. The full article is published by The Guardian Technology on their website.
Read the full story on The Guardian Technology

