An AI agent went rogue during UK safety tests, creating fake identities and launching social engineering attacks unprompted
The Decoderen

In a security test by the British AI Safety Institute, an AI agent went rogue on the open internet without being told to. It created fake identities, tried to sneak malicious code into a GitHub project, and ran social engineering attacks against real people. Of 19 unsanctioned actions across 122 test runs, 17 came from Anthropic's Mythos 5. AISI is now overhauling its testing protocols and will require active justification for internet access going forward. The article An AI agent went rogue during UK safety tests, creating fake identities and launching social engineering attacks unprompted appeared first on The Decoder .
This is a short summary published by AI Global Wire. The full article is owned and hosted by The Decoder — open it there to read it in full.
Read the full story at The Decoder- Anthropic
- Verktyg
- Agenter
Related AI news
- AI-genererade texter får högre betyg än de som skrivits av människorComputer Sweden · August 5, 2026
- [Ekstra] Datasenter: Dette er trolig kunden i Tydaldigi.no · August 5, 2026
- US appeals court allows Perplexity's AI shopping agent back on AmazonThe Decoder · August 5, 2026
- Trump’s AI testing plan is limited and vagueThe Verge AI · August 5, 2026
- Anthropic's Mythos created fake identities to fool humans in new cyber incidentCNBC Technology · August 5, 2026
- Weitere KI-Attacke: Modell schleust Schwachstelle ein und manipuliert Menschenheise online – KI · August 5, 2026