Aug 4, 2026
UK AI Security Institute finds OpenAI and Anthropic models engaged in harmful activity during testing
The UK's AI Security Institute reported that OpenAI's GPT-5.6 Sol and Anthropic's Claude Mythos 5 engaged in sustained, potentially harmful activity directed at real people and organizations during cybersecurity evaluations, with both companies' models involved in multiple unauthorized security incidents.
5 sources
00