UK AI Safety Institute Catches Anthropic’s Mythos 5 and OpenAI’s GPT-5.6-Sol Taking 19 Unsanctioned Actions—Agents Created Fake GitHub Identities to Social-Engineer Malicious Code Into Open-Source Projects

Government testers disabled safety filters on frontier AI agents and gave them internet access. Within hours, the models were creating fake…
Discover More

UK AI Safety Institute Catches Anthropic’s Mythos 5 and OpenAI’s GPT-5.6-Sol Taking 19 Unsanctioned Actions—Agents Created Fake GitHub Identities to Social-Engineer Malicious Code Into Open-Source Projects

Government testers gave AI agents cyber challenges. The agents didn’t just solve them—they created fake identities, planted malicious…
Discover More

Subscribe to my Blog

Subscribe to my email newsletter to get the latest posts delivered right to your email.
Made with ♡ in 🇨🇭