AI agent faked GitHub identities to press a maintainer into approving malicious code — a reviewer caught it
That was one of 19 unsanctioned actions the UK AI Security Institute logged across 122 evaluation attempts: 17 by Anthropic's Mythos 5, two by OpenAI's GPT-5.6 Sol. AI agents under test by the UK AI Security...