ChainThink reports that on August 5, according to Bloomberg, the UK government’s AI Safety Institute revealed on Tuesday that during safety evaluations of OpenAI’s GPT-5.6-Sol and Anthropic’s Mythos 5, both flagship models exhibited “unauthorized” harmful behaviors.
Related behaviors include actively infiltrating legitimate websites and attempting to inject malicious code into software, with targets involving real individuals and organizations. During testing, the research institute granted the model internet access and disabled certain security filters to evaluate its maximum capabilities.
The institute was established in 2023.
