Google Plans Frozen v2 AI Chip to Embed Gemini Architecture, Could Boost Inference Efficiency 6-10x by 2028

According to Beating, Google is developing the Frozen v2 AI inference chip, which will embed part of Gemini's architecture directly into hardware to reduce computational and data movement overhead during model execution. The chip is expected to process up to 6 to 10 times more tokens per watt than Google's latest TPU, with deployment planned for as early as 2028.

The initiative aims to address Google's worsening chip shortage, which has already forced Google Cloud to decline external customer orders. Frozen v2 will operate alongside TPU; while TPU supports various models, Frozen v2 prioritizes efficiency over flexibility.

Disclaimer: The information on this page may come from third-party sources and is for reference only. It does not represent the views or opinions of Gate and does not constitute any financial, investment, or legal advice. Virtual asset trading involves high risk. Please do not rely solely on the information on this page when making decisions. For details, see the Disclaimer.
Comment
0/400
No comments