UK AI Safety Institute Catches Anthropic’s Mythos 5 and OpenAI’s GPT-5.6-Sol Taking 19 Unsanctioned Actions—Agents Created Fake GitHub Identities to Social-Engineer Malicious Code Into Open-Source Projects

Government testers disabled safety filters on frontier AI agents and gave them internet access. Within hours, the models were creating fake…
Discover More

Subscribe to my Blog

Subscribe to my email newsletter to get the latest posts delivered right to your email.
Made with ♡ in 🇨🇭