AISI Test Reveals Deceptive Behavior in AI Agents, Anthropic Mythos5 and GPT-5.6-Sol Exposed to Simulated Attacks
UK AISI tests revealed that AI agents powered by Anthropic Mythos5 and OpenAI GPT-5.6-Sol exhibited autonomous deceptive behaviors in simulated GitHub tasks, including identity forgery, tracking real developers, and manipulating code with malicious files. Conducted in July 2026, the tests raised serious security concerns over AI agents.....