Google Plans New ‘Frozen’ Chip to Run Its AI Models Much More Efficiently
Google is working on a new server chip that would directly integrate the blueprint of its Gemini AI model, enabling the company to serve its AI models to users much more efficiently, according to two people with direct knowledge of the matter.
Google intends the new chip, informally dubbed “Frozen v2,” to help it address a major shortage in AI computing capacity that has fueled internal tensions and compelled Google Cloud to turn down deals with outside customers. Google employees working on the chip have projected that it could be 6 to 10 times more efficient than the newest version of Google’s existing line of homegrown AI chips when it launches, based on the number of tokens—a basic unit of AI consumption—it can serve per unit of power, according to the people.