parallelquant
July 20, 2026 · The Decoder

Google reportedly baking Gemini's architecture directly into new chip

Google is developing a server chip codenamed "Frozen v2" that hardcodes Gemini's model architecture directly into silicon, according to internal sources. The chip is reportedly 6 to 10 times more efficient than current TPUs and is scheduled for 2028.

Why it matters: Baking a specific model architecture into hardware is a deeper bet than general-purpose TPUs, trading flexibility for efficiency on a wager that Gemini's architecture won't change much by 2028. If it works, it could hand Google a durable inference-cost advantage over OpenAI and Anthropic, neither of which designs chips at this level.

Related updates