aifollow.news Search
Back Simon Willison AI 实践
Simon Willison AI 实践· · Original publication time

Anthropic red team reports GLM-5.3 control flow hijacks in binary exploitation testsMachine translation

Automatically verified and published · Generated and evidence-checked automatically; not reviewed by a human.

AI-assisted summary

Anthropic Frontier Red Team evaluated several models on 100 randomly selected tasks from its internal Binary Exploitation benchmark. GLM-5.3 developed full control flow hijacks in 4% of trials, versus 6% for Claude Mythos Preview; earlier models Claude Opus 4.6 and GLM-5.2 succeeded in none of those tasks.

The complete source text is not yet available.

Read at the original source
Found an error? Send a correction