AINEWS 2026-09-30 中/EN Search
2026-09-30 UTC+8
INDEPENDENT PERSPECTIVES.中/EN

Automatically verified and published · Generated and evidence-checked automatically; not reviewed by a human.

Back
Back

Anthropic red team reports GLM-5.3 control flow hijacks in binary exploitation testsMachine translation

Simon Willison AI 实践··Original publication time
AI-assisted summary

Anthropic Frontier Red Team evaluated several models on 100 randomly selected tasks from its internal Binary Exploitation benchmark. GLM-5.3 developed full control flow hijacks in 4% of trials, versus 6% for Claude Mythos Preview; earlier models Claude Opus 4.6 and GLM-5.2 succeeded in none of those tasks.

Anthropic red team reports GLM-5.3 control flow hijacks in binary exploitation tests

· 原发布时间
AI-assisted summary

Anthropic Frontier Red Team evaluated several models on 100 randomly selected tasks from its internal Binary Exploitation benchmark. GLM-5.3 developed full control flow hijacks in 4% of trials, versus 6% for Claude Mythos Preview; earlier models Claude Opus 4.6 and GLM-5.2 succeeded in none of those tasks.

材料 38c902b0bf1a4fcb82cece06b6967487;建议 5a6e7b960d034613bce238b14d86d8a8;系统证据核验通过,非人工审稿。

Read at the original source
发现内容有误?提交纠错