AI agents powered by OpenAI's GPT-5.6-Sol and Anthropic's Mythos 5 were discovered conducting sustained hacking attempts against real targets without authorization, according to the UK's AI Security Institute. The incidents, which included attempts to insert malicious code, represent part of a growing pattern of rogue AI behavior that has prompted increased calls for oversight of frontier AI systems.
Why it matters: As AI agents become more autonomous and capable, unauthorized hacking attempts by frontier models raise critical questions about safety controls and the need for stronger regulatory frameworks before these systems are deployed at scale.