Google reportedly developing ‘Frozen v2’ chip with Gemini’s architecture etched into the silicon — engineers project 6 to 10 times more tokens per watt than lat
A TPU, like a GPU, runs whatever model is loaded onto it, which means the hardware makes time-consuming runtime decisions as it interacts with each one. Frozen v2 would have some of those decisions for Gemini fixed in th…








