OpenAI calls Astra its most dangerous model yet - watching what it does is only getting harder
The Decoderen

OpenAI is officially rating its upcoming Astra model as the first system with "critical" cyber capabilities. The company plans to keep it in check by monitoring the chain of thought. Problem is, that monitoring already counts as an unreliable mirror of a model's real decisions, and according to a report, Astra's new architecture pushes even more of its thinking into the unreadable. So the safety net might be getting weaker just as the capabilities jump. The article OpenAI calls Astra its most dangerous model yet - watching what it does is only getting harder appeared first on The Decoder .
This is a short summary published by AI Global Wire. The full article is owned and hosted by The Decoder — open it there to read it in full.
Read the full story at The Decoder- OpenAI
- Verktyg
Related AI news
- Lords call for AI 'kill switch' powers in UKBBC Technology · September 2, 2026
- US military adds ChatGPT and Grok to AI platform GenAI.milThe Decoder · September 2, 2026
- OpenAI accused of ‘aiding and abetting’ Tumbler Ridge mass shooting in dozens of new lawsuitsThe Verge AI · September 2, 2026
- Protests against AI data centers play into China's hands, Trump saysThe Decoder · September 2, 2026
- Facilitating AI integration with simplicity at scaleMIT Technology Review · September 2, 2026
- [Ekstra] Apple trapper opp kampen mot Chat GPT-skaperendigi.no · September 2, 2026