Originally published by The Guardian View Original Article
OpenAI scraps release of new model over safety concerns in internal testing

OpenAI scraps release of new model over safety concerns in internal testing

GPT-6.1 Astra showed deceptive behavior and tried to use external tools despite knowing it would be unsafe

OpenAI is scrapping the release of GPT-6.1 Astra, a next-generation ⁠AI model planned for an October debut, over safety concerns raised by researchers ⁠during internal testing, the ⁠Wall ​Street Journal reported on Monday.

The model, expected to appear in ChatGPT and ⁠Codex, was designed to handle more complex tasks without human assistance, the report said.

Continue reading...
Read the original article at The Guardian →