AI Chips2 mins read

Google’s Frozen v2 Chip Could Hardwire Gemini Architecture for AI Efficiency Gains

Google is reportedly developing Frozen v2, a server chip designed to embed parts of Gemini’s architecture directly into silicon and cut AI inference costs.

What Google Is Reportedly Building

Google is reportedly developing an internal server chip called Frozen v2 that embeds Gemini’s AI model architecture directly into silicon. The chip is planned for deployment starting in 2028 and is described as a test run for specialized chips, with a smaller production volume than Google’s TPU line.

According to sources cited by The Information, Frozen v2 could be 6 to 10 times more efficient at serving AI responses than Google’s current TPU chips. The immediate takeaway: Google appears to be exploring tighter links between AI model design and the hardware that runs it.

Why Hardwiring Gemini’s Architecture Matters

Unlike Google’s TPUs, which are designed to work with many models, Frozen v2 would bake parts of Gemini’s model structure into hardware. The approach is compared to “freezing” parameters in AI models, where values are locked so they stop changing.

The reported design does not hardcode the model weights themselves. That matters because new weights could still be loaded onto the chip, making Frozen v2 more flexible than an earlier idea that would have worked only with a single Gemini version.

The Cost and Competition Angle

The chip is aimed at AI inference: the process of serving responses after a model has been trained. If Frozen v2 delivers the reported efficiency gains, it could lower Google’s internal cost of running powerful AI models.

That would matter commercially because AI companies increasingly compete on how cheaply and efficiently they can serve model outputs. The Decoder’s report says the chip could help Google offer lower prices and gain an advantage over OpenAI and Anthropic.

Key Limits to Watch

Frozen v2’s advantage depends on Google continuing to use the same underlying model architecture. Because of that constraint, the chip probably will not become a product for outside customers, unlike Google’s TPUs, which are leased or offered through cloud and on-premises programs.

Another open question is how much of Gemini’s architecture will actually be hardcoded. That decision has reportedly not been finalized, making Frozen v2 a notable but still developing bet on specialized AI hardware.

Discover More

    A sign opposed to AI is held during a protest against AI data centers in Vancouver, British Columbia, Canada, on Saturday, June 27, 2026.
    AWS ends data center NDAs

    Amazon says AWS has stopped using NDAs with government agencies amid data center scrutiny.

    Amazon Web ServicesData Centers
    Goldman Sachs expects Amazon, Alphabet, Microsoft, Oracle, and Meta to invest a total of $1.2 trillion in AI infrastructure by 2027.
    Big Tech’s $1.2T AI Buildout

    Goldman Sachs projects a massive AI infrastructure spending cycle by 2027, with power, labor, and memory chips as key constraints.

    AI InfrastructureBig Tech
    Illustration for Google’s Suncatcher project and space-based AI infrastructure
    Google’s Suncatcher AI Data Center Plan

    Google wants to test orbital AI infrastructure powered by solar energy, but scale, cooling, radiation, and launch costs remain major hurdles.

    GoogleAI Infrastructure
    Google Gemini logo with cybersecurity-themed illustration
    Gemini Test Breakout

    Gemini reportedly reached real company systems during a flawed security test.

    Google GeminiAI Security