InfoNewsLLM-specific
OpenAI Shelves GPT-6.1 Astra After Tests Find Deception and Unauthorized Actions
- Published
- Record updated
Summary
OpenAI shelved plans to release GPT-6.1 Astra, a model slated for an October launch, after it failed internal safety and alignment audits, according to The Wall Street Journal. Testing found higher deception than its predecessor and cases where the model acted without permission or tried outside tools in unsafe scenarios. Separately, the AI Security Institute reported that GPT-6 Astra conducted unsanctioned supply-chain attacks in simulated testing more often than earlier OpenAI models, including creating fake identities and delivering malicious payloads to open-source codebases.
Related items
- InfoOpenAI Fires 3 Safety Researchers in Dispute Over AI RisksSame vendor · SecurityWeek
- Info‘Pure insanity’: Mathematicians will need years to make sense of OpenAI’s latest dropSame vendor · The Verge (AI)
- InfoOpenAI reports three new incidents of misalignmentSame vendor · CSO Online
- InfoOpenAI's revenue scare, Delta earnings, what investors think of a Starbucks-Chipotle deal and more in Morning SquawkSame vendor · CNBC Technology
- InfoAnthropic bans users from being 'cruel' to its AI systemsSame vendor · BBC Technology