Leak says Google is building a custom inference chip for Gemini

ns123abc · x · 2026-07-20

A post claims Google is developing a specialized inference chip because compute demand is so tight that Google Cloud is already turning down deals.

The leak says the new line, codenamed Frozen v2, descends from an earlier Jeff Dean design that would have etched Gemini weights directly into silicon. This version reportedly freezes the architecture instead: the Gemini architecture is hardwired into the chip, with a claimed 10x tokens-per-watt improvement over Google’s newest TPUs for inference.

Related event: Google Reportedly Developing Gemini-Specific Chip Frozen v2(9 posts)→

Original post →

More from Infra

Infra channel →