Fetching from the wire…
Public story · 2026-07-22 · high
CryptanalysisBench had five frontier models break 65 to 86 percent of already-broken ciphers across 191 tasks.
Why now: The arXiv paper is part of the July 22 coverage of frontier-model capability claims, one of the few dated data points behind the AI-does-real-research argument.
Five frontier models cracked 6 to 12 previously unbroken cryptographic schemes at full strength, per a benchmark called CryptanalysisBench. Two of those cracks were new to the field: a key-recovery vulnerability in SpoC AEAD and an error in KINDI's security proof. For anyone maintaining or reviewing a custom cipher, that's the review process failing at its one job.
The benchmark ran 191 tasks across six families of cryptographic primitives, posted to arXiv as 2607.18538. It broke 65 to 86 percent of ciphers already known to be broken. Scaled-down versions of the hardest problems fell 24 to 61 times, tested across Claude Opus 4.8, Sonnet 5, Mythos 5, GPT 5.5, and GLM 5.2.
Yes, but reproducing a textbook attack is pattern matching, and that's most of the 65 to 86 percent.
The arXiv post doesn't say whether SpoC AEAD's or KINDI's maintainers have been notified, or whether either flaw holds up outside the benchmark's controlled conditions.
Each link below shares sources, entities, or timing with this story.
OpenAI released Frontier / Shared entities / Shared topic / Earlier coverage
Linked by a graph relationship (OpenAI released Frontier); both cover Claude Opus, Frontier, GLM, GPT; overlapping topics (claude, model).
OpenAI released Frontier / Shared entities / Earlier coverage / Tension
Linked by a graph relationship (OpenAI released Frontier); both cover Claude Opus, GLM, GPT, LLM; earlier Claude Opus coverage from 2026-04-20.
OpenAI released Frontier / Shared entities / Shared topic / Earlier coverage
Linked by a graph relationship (OpenAI released Frontier); both cover Claude Opus, GPT, Models; overlapping topics (claude, model).
Linked by a graph relationship (OpenAI released Frontier); both cover GLM, LLM, Models; overlapping topics (claude, model).
OpenAI released Frontier / Shared entities / Same source domain / Shared topic / Earlier coverage
Linked by a graph relationship (OpenAI released Frontier); both cover Claude Opus, GPT; reported by the same outlet (arxiv.org).
OpenAI released Frontier / Shared entities / Earlier coverage / Tension
Linked by a graph relationship (OpenAI released Frontier); both cover Claude Opus, Frontier, GPT; earlier Claude Opus coverage from 2026-04-25.
OpenAI released Frontier / Shared entities / Earlier coverage
Linked by a graph relationship (OpenAI released Frontier); both cover Frontier, GPT, Mythos; earlier Frontier coverage from 2026-07-10.
Linked by a graph relationship (OpenAI released Frontier); both cover Frontier, GPT, Mythos; earlier Frontier coverage from 2026-06-29.