Reddit
The Information: Google's 'Frozen v2' Chip Etches Gemini's Architecture Into Silicon for 6–10x Tokens per Watt
Reported July 20, Google is developing Frozen v2, a chip that hardwires parts of Gemini's neural architecture directly into the circuitry so less computation and data movement is needed per response — engineers project 6 to 10 times more tokens per watt than current TPUs, targeted as early as 2028. The idea reportedly originated with DeepMind chief scientist Jeff Dean, and Google frames it as a new chip family alongside TPUs rather than a replacement, motivated by an internal compute shortage severe enough that Google Cloud has been turning customers away. The strategic bet worth noting: baking architecture into silicon only pays off if transformer-style model architecture has stopped changing.
↳ Follow the thread