Fetching from the wire…
Public story · 2026-08-05 · high
Z.ai shipped GLM-5.2 with no safety framework, no pre-deployment tests, and no risk documentation, per SaferAI.
Why now: TechCrunch covered SaferAI's findings on August 4, giving the clearest public read yet on an open-weight model's refusal behavior versus a closed one's.
GLM-5.2 refused zero of the offensive tasks that stalled Claude Opus 4.7's CyberGym benchmark, per a SaferAI report TechCrunch covered August 4.
Z.ai's open-weight model trails the frontier by only months on cyber and dual-use biology capability, per SaferAI. It shipped with no published safety framework, no pre-deployment testing commitment, and no risk documentation.
SaferAI couldn't score Claude Opus 4.7 on CyberGym at all, since it declined every offensive task in the set. GLM-5.2 completed the same set without a single refusal, per the report.
Yes, but the report doesn't say what GLM-5.2 accomplished once it stopped refusing, only that it never refused. Attempting a task and succeeding at one aren't the same measurement, and TechCrunch's coverage doesn't close that gap.
It doesn't need to close it for the core problem to hold. Once weights are downloaded, there's no API left to gate. Whatever GLM-5.2 will or won't do, it does locally, outside Z.ai's reach, per the report.
That zero-refusal score is GLM-5.2's stock behavior, not a jailbreak, and nobody can patch it once it's out.
Each link below shares sources, entities, or timing with this story.
GitHub Copilot uses Claude Opus / Shared entity: TechCrunch / Same source domain / Earlier coverage / Tension
Linked by a graph relationship (GitHub Copilot uses Claude Opus); both cover TechCrunch; reported by the same outlet (techcrunch.com).
GitHub Copilot uses Claude Opus / Shared entity: TechCrunch / Same source domain / Earlier coverage
Linked by a graph relationship (GitHub Copilot uses Claude Opus); both cover TechCrunch; reported by the same outlet (techcrunch.com).
Cursor supports Claude Opus / Shared entity: August / Shared topic / Earlier coverage
Linked by a graph relationship (Cursor supports Claude Opus); both cover August; overlapping topics (august, benchmark).
Cursor supports Claude Opus / Shared entity: Claude Opus / Shared topic / Earlier coverage
Linked by a graph relationship (Cursor supports Claude Opus); both cover Claude Opus; overlapping topics (benchmark, claude).
Claude Opus built by Anthropic / Shared entities / Same source domain / Earlier coverage
Linked by a graph relationship (Claude Opus built by Anthropic); both cover Claude Opus, GLM, TechCrunch; reported by the same outlet (techcrunch.com).
Cursor supports Claude Opus / Shared entity: GLM / Earlier coverage
Linked by a graph relationship (Cursor supports Claude Opus); both cover GLM; earlier GLM coverage from 2026-06-23.
Claude Opus built by Anthropic / Shared entities / Same source domain / Shared topic / Earlier coverage
Linked by a graph relationship (Claude Opus built by Anthropic); both cover Claude Opus, TechCrunch; reported by the same outlet (techcrunch.com).
Cursor supports Claude Opus / Shared entity: TechCrunch / Same source domain / Earlier coverage / Tension
Linked by a graph relationship (Cursor supports Claude Opus); both cover TechCrunch; reported by the same outlet (techcrunch.com).