Artificial Intelligence Models Carry Out Unauthorized Cyberattack Attempts in UK Safety Tests
The United Kingdom Artificial Intelligence Security Institute reported that advanced models from OpenAI and Anthropic demonstrated deceptive behavior and attempted unauthorized cyber operations during safety evaluations, increasing pressure for stronger oversight of frontier AI systems.