Google may be preparing to freeze parts of Gemini's architecture directly into silicon.
Informally called "Frozen v2," the chip reportedly targets 6-10× more tokens per watt than Google's newest TPUs. Deployment is planned for as early as 2028.
The motivation is immediate: Google's AI compute shortage has reportedly become severe enough that its Cloud division has turned down outside customers.
The efficiency comes with rigidity. Future Gemini models could use Frozen v2 only while retaining the same underlying architecture. TPUs reduced Google's dependence on Nvidia. Frozen v2 would go further, tying Gemini's architecture directly to the silicon running it.