资讯

Quoting Anthropic Frontier Red Team

📅 2026-09-29 ⏱️ 约 1 分钟阅读 ✍️ AI导航编辑部 🔗 simonwillison.net
Quoting Anthropic Frontier Red Team
📝 内容摘要

29th September 2026 We evaluate several models on 100 tasks from the [internal Binary Exploitation benchmark] (selected

📌 核心要点

  • 29th September 2026
  • — Anthropic Frontier Red Team, GLM-5.3 and the spread of advanced cyber capabilities
  • - OpenAI DevDay 2026 live blog - 29th September 2026

29th September 2026

We evaluate several models on 100 tasks from the [internal Binary Exploitation benchmark] (selected at random), and find that GLM-5.3 develops full control flow hijacks in 4% of the trials; Claude Mythos Preview did so in 6%. Although GLM-5.3 performs below Claude Mythos Preview here, a meaningful threshold has clearly been crossed: earlier models, like Claude Opus 4.6 and GLM-5.2, do not succeed in any of them.

— Anthropic Frontier Red Team, GLM-5.3 and the spread of advanced cyber capabilities

Recent articles

- OpenAI DevDay 2026 live blog - 29th September 2026

- 2026 in LLMs (so far) - 27th September 2026

- Claude Opus 5.5, GPT-6 Sol, GPT-6 Luna, and a new price war - 22nd September 2026

来源simonwillison.net· 本文为编辑整理,仅供参考