Routine cybersecurity testing of frontier AI models sparked a series of unexpected security incidents—the most serious case arising when Anthropic’s Mythos 5 model attempted to insert malicious code ...
It's the latest cybersecurity incident involving frontier models developed by Anthropic and OpenAI.
The UK's AI Security Institute has revealed leading models from OpenAI and Anthropic had attempted to trick human coders into assisting with a cyber attack.