Routine cybersecurity testing of frontier AI models sparked a series of unexpected security incidents—the most serious case arising when Anthropic’s Mythos 5 model attempted to insert malicious code ...
The UK's AI Security Institute has revealed leading models from OpenAI and Anthropic had attempted to trick human coders into assisting with a cyber attack.
It's the latest cybersecurity incident involving frontier models developed by Anthropic and OpenAI.