Fetching from the wire…
Infra2026-06-28 · source-backed
June local-inference benchmarks across the 128GB class put NVIDIA's DGX Spark ($4k), AMD's Strix Halo / Ryzen AI Max+ 395 ($2 to 3k), and the M5 Max 128GB (~$5k) head to head (Hardware Corner). Prompt processing favors CUDA hard. But token generation lands at a surprisingly close 34 to 38 tok/s on 120B models. Strix Halo's load-bearing number is ~180 GB/s real usable bandwidth, and the community runtime is Vulkan via llama.cpp, not ROCm. Translation: prompt-heavy agentic loops want a CUDA box, but chat-style generation on the cheaper AMD option is genuinely competitive. Buy for your actual workload shape, not the headline number.
Each link below shares sources, entities, or timing with this story.
NVIDIA invested in OpenAI / Shared entities / Shared topic / What happened next
Linked by a graph relationship (NVIDIA invested in OpenAI); both cover AMD, CUDA, M5 Max, NVIDIA; overlapping topics (spark, token).
NVIDIA invested in OpenAI / Shared entities / What happened next / Tension
Linked by a graph relationship (NVIDIA invested in OpenAI); both cover AMD, Nvidia, ROCm; picks up the AMD thread on 2026-07-24.
NVIDIA released Blackwell / Shared entities / Earlier coverage / Downstream implication
Linked by a graph relationship (NVIDIA released Blackwell); both cover AMD, CUDA, NVIDIA; earlier AMD coverage from 2026-04-10.
NVIDIA partners with Cisco / Shared entities / What happened next
Linked by a graph relationship (NVIDIA partners with Cisco); both cover AMD, CUDA, NVIDIA; picks up the AMD thread on 2026-07-27.
NVIDIA released Blackwell / Shared entities / Shared topic / Earlier coverage
Linked by a graph relationship (NVIDIA released Blackwell); both cover DGX Spark, NVIDIA; overlapping topics (benchmark, number, token).
NVIDIA released DGX Spark / Shared entities / Shared topic / What happened next
Linked by a graph relationship (NVIDIA released DGX Spark); both cover AMD, Nvidia; overlapping topics (agentic, bandwidth, benchmark).
NVIDIA released DGX Spark / Shared entities / Shared topic / Earlier coverage
Linked by a graph relationship (NVIDIA released DGX Spark); both cover AMD, NVIDIA; overlapping topics (bandwidth, generation).
NVIDIA invested in OpenAI / Shared entities / What happened next / Tension
Linked by a graph relationship (NVIDIA invested in OpenAI); both cover AMD, Nvidia; picks up the AMD thread on 2026-07-25.