We evaluate several models on 100 tasks from the [internal Binary Exploitation benchmark] (selected at random), and find that GLM-5.3 develops full control flow hijacks in 4% of the trials; Claude Mythos Preview did so in 6%.
Quoting Anthropic Frontier Red Team
About this summary. This is a short, independently written summary of an article first published by Simon Willison. Cyber Security News did not report or verify the underlying story. Read the original: https://simonwillison.net/2026/Sep/29/anthropic-frontier-red-team/
Source attribution: headline and facts are from Simon Willison (simonwillison.net). Summary method: excerpt of the source description. See our source attribution policy.






