Fetching from the wire…
Infra2026-07-24 · source-backed
The Chips and Cheese teardown details eight Accelerator Complex Dies on TSMC N2 plus two I/O dies and two fabric-and-cache dies on N3, twelve HBM4 stacks on 2048-bit buses, 192 channels, 256 Work Group Processors at 2.4GHz for up to 40.26 PFLOPS of OCP MXFP4. The Helios rack pairs four MI455X with one 96-core EPYC 9006 SP7 per compute tray. Here's the number builders can act on: 432GB per GPU means a 1.4TB MXFP4 model fits in four GPUs instead of a rack.
Each link below shares sources, entities, or timing with this story.
AMD unveiled its first rack-scale system to directly contest Nvidia at the rack level, with engineering samples in H2 2026 and mass production targeted Q2 2027. Microsoft joins Meta, OpenAI and Oracle as customers; Meta plans 1 gigawatt of Helios racks by year-end against a lo...
Four times MI355X on peak MXFP4, 20.13 petaflops each at MXFP6 and MXFP8, 315 teraflops vector FP16 and matrix FP32, and 23.3 TB/s of memory bandwidth per GPU, up from 288GB of HBM3E (ServeTheHome). Eight accelerator complex dies on TSMC N2 with fabric, cache and I/O dies on N...
PR #19378 landed in llama.cpp this week, and I think most people are underselling what it means. Backend-agnostic tensor parallelism via --split-mode tensor makes multi-GPU inference work across AMD, Intel, and Apple Silicon. Not just CUDA. Everything. For context: llama.cpp h...
TrendForce reporting via Tom's Hardware says prototype variants run 192GB or 256GB, down from the announced 288GB, with some configs using fewer than 16 stacks and substituting HBM4 for HBM4E. Driver is tightening HBM supply across SK hynix, Samsung, and Micron. At GTC 2025 co...
For about a year, "run your agent locally" meant accepting a model that couldn't reliably call a tool twice in a row. That excuse is gone. Meta Superintelligence Labs published Muse Glimmer today: a 29.6B dense causal transformer, 52 layers, 6,656 hidden dim, with a ~1.8B ViT-...
Huang used his inaugural X post on July 24 to publish "Open Weights and American AI Leadership," a three-page letter on Nvidia's own servers signed by 25 companies including Meta, Microsoft, IBM, Mistral, Mozilla, Hugging Face, a16z, Palantir and the Linux Foundation. Within a...
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.