Fetching from the wire…
Public story · 2026-08-25 · high
The Vera Rubin NVL72 numbers come from recorded coding sessions with tool calls and sub-agent spawning intact, not single-turn benchmarks.
Why now: The comparison surfaced in coverage dated August 25, with the numbers still unverified by outside review.
Vera Rubin NVL72 reaches 30x the throughput per megawatt of the current GB300 NVL72, per Nvidia's Vera Rubin efficiency post. The company also claims a 35x drop in cost per million tokens. That number matters most for teams sizing agent deployments, where a session's context grows across tool calls and sub-agent spawns before finishing a task.
Nvidia measured both figures on SemiAnalysis's AgentX benchmark, built from recorded agentic coding sessions rather than single-turn prompts. DeepSeek V4 Pro and Qwen3.5 were among the models tested. AgentX scores a full session. It counts context growth, tool calls, and sub-agent spawning, not a one-shot question and answer.
The results are vendor-reported and still waiting on SemiAnalysis's own review. There's no GA date for Vera Rubin NVL72, and Nvidia's post doesn't say when that review will publish.
The benchmark choice signals a shift in how chip vendors market inference economics, from single-turn token counts to full agent sessions. Watch for SemiAnalysis to publish its own AgentX numbers on the same hardware, independent of Nvidia's release, before treating 30x as settled.
Each link below shares sources, entities, or timing with this story.
Vera Rubin NVL72 competes with Blackwell / Shared entities / Same source domain / Shared topic / Earlier coverage
Linked by a graph relationship (Vera Rubin NVL72 competes with Blackwell); both cover GB300 NVL72, NVIDIA, Qwen3; reported by the same outlet (blogs.nvidia.com).
NVIDIA uses Claude Code / Shared entities / Shared topic / Earlier coverage
Linked by a graph relationship (NVIDIA uses Claude Code); both cover DeepSeek V4 Pro, Qwen3; overlapping topics (agent, agentic, coding, token).
NVIDIA released Vera Rubin NVL72 / Shared entities / Shared topic / Earlier coverage
Linked by a graph relationship (NVIDIA released Vera Rubin NVL72); both cover NVIDIA, Qwen3; overlapping topics (against, agent, token).
NVIDIA released Vera Rubin NVL72 / Shared entity: NVIDIA / Same source domain / Shared topic / Earlier coverage / Tension
Linked by a graph relationship (NVIDIA released Vera Rubin NVL72); both cover NVIDIA; reported by the same outlet (blogs.nvidia.com).
NVIDIA released Vera Rubin NVL72 / Shared entity: NVIDIA / Shared topic / Earlier coverage / Tension
Linked by a graph relationship (NVIDIA released Vera Rubin NVL72); both cover NVIDIA; overlapping topics (against, agent, call, coding).
Alibaba partners with NVIDIA / Shared entity: Qwen3 / Shared topic / Earlier coverage / Tension
Linked by a graph relationship (Alibaba partners with NVIDIA); both cover Qwen3; overlapping topics (agent, agentic, benchmark, coding).
NVIDIA released Vera Rubin NVL72 / Shared entity: NVIDIA / Same source domain / Shared topic / Earlier coverage
Linked by a graph relationship (NVIDIA released Vera Rubin NVL72); both cover NVIDIA; reported by the same outlet (blogs.nvidia.com).
NVIDIA released Vera Rubin NVL72 / Shared entities / Earlier coverage
Linked by a graph relationship (NVIDIA released Vera Rubin NVL72); both cover GB300 NVL72, NVIDIA, Qwen3; earlier GB300 NVL72 coverage from 2026-08-13.