Codex Reset
GitHub

QQ·微信群

CODEX / SIGNAL STUDIO

Your Codex, in focus.

GLM-5.3 scores 12% on ExploitBench, near Claude Mythos at 14%

DeepLearning.AI

[DeepLearning.AI’s weekly letter](https://hubs.la/Q04z3MJH0) says an open model, GLM-5.3, nearly matches Claude Mythos on cybersecurity tasks. According to Anthropic, GLM-5.3 solves 12% of ExploitBench exploit tasks and Claude Mythos solves 14%.

The letter says $20.40 in tokens found a recent security exploit in Google Chrome. Andrew Ng writes that concern about these capabilities in the wrong hands should be treated as an engineering problem, and that defenders hold the long-term edge.