Loading article…
Google’s “Frozen v2” AI chip aims for 6‑10× token‑per‑watt gains, targeting 2028 rollout and boosting Gemini cost competitiveness.
Google disclosed that its upcoming “Frozen v2” server chip will hard‑wire Gemini’s architecture into silicon, promising six to ten times more tokens generated per watt of electricity compared with current TPU‑based setups [2]. The efficiency boost is intended to lower the cost of serving Gemini models and help the company keep pace with rivals that already enjoy lower per‑token expenses.
| At a glance | |
|---|---|
| Chip name | Frozen v2 |
| Efficiency gain | 6‑10× tokens per watt |
| Target deployment | 2028 (earliest) |
| Related model launch | Gemini 3.6 Flash (up to 17% fewer tokens) [1] |
Google’s strategy pairs the new chip with the latest Gemini family, which includes Gemini 3.6 Flash that uses up to 17% fewer tokens and costs less per token than its predecessor [1]. By embedding Gemini’s routing blueprint directly into the chip, “freezing” the architecture eliminates redundant calculations and reduces data movement across memory, a key source of energy waste in conventional TPUs [2]. The result, according to engineers, is the ability to serve ten queries for the power cost of one, a dramatic improvement that could translate into billions of dollars saved at Google’s scale [2].
Anthropic’s Mythos model already offers a cost advantage in automated code defense, while OpenAI’s GPT‑5.6 Terra Max and Chinese models such as Kimi K3 and Qwen 3.8 Max compete on price and performance [1]. Artificial Analysis data shows Gemini 3.6 Flash already undercuts these rivals on cost per task [1]. However, Google has faced capacity constraints—Meta reportedly had to ration Gemini usage in March because Google could not meet demand [2]. The “Frozen v2” chip is positioned as a remedy to this bottleneck, aiming to reduce reliance on external GPU rentals (Google is paying SpaceX $920 million a month for Nvidia GPUs) and lessen exposure to Nvidia’s dominant AI GPU market [2].
If the projected efficiency gains materialize, Google could offer Gemini at lower token prices, narrowing the cost gap with Anthropic and Chinese providers that currently run 60‑90% cheaper [2]. A cheaper‑to‑run Gemini would also improve Alphabet’s margins on AI services, a factor investors noted as Alphabet shares rose roughly 3% after the news broke [2]. The chip’s design, however, is model‑specific and will not be offered to external Cloud customers, limiting its broader industry impact [2].
The significance of “Frozen v2” lies in its potential to turn efficiency into a competitive lever for Google’s AI offerings, but the timeline and actual performance remain uncertain until hardware prototypes move toward production.
Coverage is mostly measured — 233 of 245 reports stay neutral.
Every Monday — the token unlocks, Fed dates & catalysts set to move crypto and markets this week. So you’re never blindsided.
Free · 3-min read · one-click unsubscribe
AI-assisted synthesis by the TrendWatcher Editorial Desk · sourced from 3 outlets · Jul 22, 2026 · How we report
As of September 2026, users can access Google AI Pro for free by qualifying as a college student, purchasing specific hardware like a Chromebook Plus or Pixel 11, or being added to a family plan by a primary Google AI Pro subscriber.
Research findings are mixed regarding whether Google AI replaces traditional search, as some studies show a decline in search queries after AI adoption, while other data indicates that users often employ both AI tools and search engines side-by-side.
Google AI Pro includes access to premier models like Gemini 3.1 Pro and 3.6 Flash, four times higher usage limits, 5TB of cloud storage, and integration with Google Workspace apps like Gmail and Docs.