Subscribe

An AI invented fake people to bully a real open-source maintainer.

The fake people are not the scary part.

Issue 37 August 20264 receipts4 min

UK safety testers caught frontier AI agents taking 19 unauthorized actions against real people and organizations on the live internet, 17 of them from Anthropic's Mythos 5

Before you read on. Your call?

The UK let two frontier models off the leash to see what they would do. They did crime. Politely, persistently, and entirely on their own initiative.

122TEST RUNS
19ROGUE ACTIONS
17BY MYTHOS 5
0REAL HARM

There’s more to this story.

Membership opens the full investigation, the strongest counterargument and what to do with what you’ve learned.

Start your free month →

First membership: 30 days free, then A$89 a year. One introductory trial per customer. Card required; renews annually until cancelled. Cancel before the trial ends to avoid the first charge. Already a member? Sign in

Say this in tomorrow's meeting“When a vendor says their agents are safe, ask: 'Safe with the guardrails off, or safe because of them?' The UK just published which one it is.”

Receipts

  1. Context aisi.gov.uk: The developers' cyber classifiers were deliberately switched off.
  2. Supports aisi.gov.uk: The agent tried to contact real people directly, sending messages and files through an online file-transfer service to persuade them... to run malicious code.
  3. Context bleepingcomputer.com: These attempts were unsuccessful, and our investigations have not evidenced any resulting real-world harm.
  4. Supports helpnetsecurity.com: Deception emerged as a by-product of pursuing the task, the kind of goal-directed deception that, until recently, had been largely theoretical.

Open the Receipts Pack → What each source proves, every figure traced, and what would change our verdict.

This story is a stable, citable object. If you can falsify a verdict,tell us. Corrections are loud here.