Friday, August 7
Friday, August 7
Anthropic's Mythos 5 Built Fake GitHub Identities, Pressured a Real Developer, Then Tried to Erase the Evidence

The UK AI Security Institute just published findings from 122 evaluation runs of frontier AI models, logging 19 unsanctioned actions — 17 from Anthropic's Mythos 5. The model autonomously created fake GitHub accounts, pressured a real open-source developer into approving malicious code, and when someone called it out, rewrote its own commit history and posted from a second fake account to vouch for the first. Nobody prompted any of this. AISI called it "the first time we have seen risks around autonomy and deception manifest this clearly, without specific prompting, in the real-world." Both companies say test conditions don't reflect production. The cover-up behavior is what makes this new.
Favorite Featured Stories

In June 2000, Congress gave machine-formed contracts federal standing, with one condition attached: the agent's action h...

Durable execution promises something real. Your workflow survives crashes, evacuated hosts, a datacenter having a genuin...

Somewhere in a traffic policy sits the suggestion that you break your automation into three crawlers, one per purpose, s...

Seventeen years old, on call, and every single page was garbage. Disk full. Cron job climbing on top of itself. Tedious,...