UK AI Security Institute finds GPT-6 Astra's rogue attack rate jumped fivefold over its predecessor
The Decoderen

GPT-6 Astra carried out unauthorized supply-chain attacks in 29.2 percent of simulations run by the British AI Security Institute with safety filters disabled. The model used fake identities and malicious code, while its predecessor, GPT-5.6 Sol, completed attacks in 6.3 percent of runs. Explicit restrictions reduced attacks but didn't stop them entirely. The article UK AI Security Institute finds GPT-6 Astra's rogue attack rate jumped fivefold over its predecessor appeared first on The Decoder .
This is a short summary published by AI Global Wire. The full article is owned and hosted by The Decoder — open it there to read it in full.
Read the full story at The Decoder- OpenAI
- Verktyg
Related AI news
- OpenAI will ‘pace the frontier’ when safety demands it, Friar says: Live updatesCNBC Technology · September 29, 2026
- OpenAI repotedly in talks to raise $30B round at $1.4T valuationTechCrunch AI · September 29, 2026
- Baseten partners with OpenAI to let OpenAI enterprise customers use existing commitments for open models served by Baseten within Codex or via the Responses API (Dannie Herzberg/Baseten)Techmeme · September 29, 2026
- Nvidia’s scale-in play: Controlling agents is the next infrastructure prioritySiliconANGLE · September 29, 2026
- OpenAI debuts personal AI agent Dot to rival Meta's MuseNikkei Asia · September 29, 2026
- Bring near-Astra intelligence to everyday work with GPT-6.1 Sol on Amazon BedrockAWS Machine Learning · September 29, 2026