(Image credit: Google) Google is developing a server chip, informally dubbed "Frozen v2," that would etch part of its Gemini model's architecture directly into the silicon, according to a report published Monday by The Information , citing two people with direct knowledge of the matter. Engineers on the project have projected that the chip could serve six to ten times more tokens per unit of power than the newest generation of Google's TPUs, with deployment targeted for as soon as 2028. The two sources said the project is partly a response to an AI compute shortage severe enough that Google Cloud has turned down deals with outside customers. Go deeper with TH Premium: AI and data centers A TPU, like a GPU, runs whatever model is loaded onto it, which means the hardware makes time-consuming runtime decisions as it interacts with each one.…