| Mobile Web

Google develops Frozen v2 AI inference chip, up to 10 times more efficient than TPU

Google is developing a new server chip that directly integrates the design of its Gemini AI model, The Information reported on July 21, citing two people familiar with internal matters. The chip, called Frozen v2 inside Google, is aimed at addressing an AI computing power shortage that has fueled internal conflict and led Google Cloud to turn down external customer contracts. Employees involved expect Frozen v2 to be 6 to 10 times more efficient than the latest TPU by tokens processed per unit of power. Google plans to deploy it as early as 2028.