Google Developing 'Frozen v2' Chip for Gemini Efficiency Gains
Alphabet's Google is developing a new server chip designed to make its Gemini AI models more efficient, according to Tom's Hardware. The processor, known internally as 'Frozen v2', could deliver between six and ten times more tokens per unit of power than Google's current tensor processing units.
The chip would embed parts of Gemini's architecture directly into the silicon to reduce computational steps and data movement. Engineers aim to deploy it by 2028, focusing on improving inference performance while reducing power usage. The project originates from research led by DeepMind chief scientist Jeff Dean.
Unlike earlier versions that proposed baking model weights into silicon, Frozen v2 will retain flexibility by fixing only parts of the model architecture, allowing updates across future Gemini releases. The design remains experimental and will complement Google's existing TPU hardware instead of replacing it.
This chip development follows broader efforts by major AI companies to create specialized hardware that reduces dependence on suppliers such as Nvidia and addresses capacity constraints in AI infrastructure.
We hope you enjoyed this article.
Consider subscribing to one of our newsletters like Silicon Brief or Daily AI Brief.
Also, consider following us on social media:
More from: Data Centers
Subscribe to Silicon Brief
Weekly coverage of AI hardware developments including chips, GPUs, cloud platforms, and data center technology.
Whitepaper
AI and the Law: Discussion Paper
This discussion paper explores the intersection of artificial intelligence and legal frameworks, addressing potential legal challenges posed by AI's autonomy, adaptiveness, and opacity. It examines issues such as liability gaps, causation, and the potential for granting AI systems legal personality, aiming to foster further discussion on AI's impact on law reform.
Read more