Google develops Frozen v2 Gemini AI chip
Google is developing a specialized AI server chip codenamed "Frozen v2" that embeds Gemini model architectures directly into silicon. Slated for deployment by 2028, the chip is expected to yield 6 to 10 times better energy efficiency than existing TPUs.
Hardwiring model architectures is a bold bet on hardware-software co-design, offering massive efficiency gains at the cost of long-term flexibility. Direct hardware implementation of model layers eliminates memory access overheads, which are the primary source of energy consumption in inference. While an order-of-magnitude efficiency improvement is crucial for scaling Gemini services sustainably, designing a chip for deployment in 2028 risks committing to design architectures that could become obsolete before launch.
DISCOVERED
13h ago
2026-07-20
PUBLISHED
14h ago
2026-07-20
RELEVANCE
AUTHOR
mark_k