Michael Burry says markets should crash hard to stop OpenAI and Anthropic from going public. “For the benefit of humanity.”

OpenAI scraps release of new AI model after it showed ‘higher levels of deception’

OpenAI has scrapped plans to release its new AI model amid safety concerns, after it showed “higher levels of deception” than its predecessor.

Saachi Jain, the company’s safety chief, revealed in an interview with the Wall Street Journal that GPT-6.1 Astra, which had been scheduled for release in October, had at times failed to ​accurately disclose actions it had or ​had not taken.

Ms Jain told the outlet the new model, which was set to appear in ChatGPT and Codex and designed to handle more complex ‌tasks without human assistance, ​fell short of OpenAI’s standards in alignment tests, which assess whether a system follows human intent.

She added that GPT-6.1 Astra also had problems with “scope authorisation”, proceeding with tasks ⁠without requesting user permission and attempting to use external tools or services when doing so ⁠could be unsafe.

Got a news tip or correction? Let us know

If you got something out of this, please chip in to keep this site running, or subscribe to go ad-free.

0 views