हिंदी में पढ़ें —JantaScope हिंदी
AI NEWS

Google Reportedly Developing ‘Frozen v2’ AI Server Chip to Improve Gemini Performance and Power Efficiency.

Google is reportedly working on a new in-house AI server chip, informally known as "Frozen v2," to make its Gemini artificial intelligence models run more efficiently. According to reports, the specialized processor would integrate elements of Gemini directly into the hardware, potentially delivering significantly higher performance while consuming less power. The project is still under development and is expected to complement—not replace—Google's existing Tensor Processing Units (TPUs).

Google Reportedly Developing ‘Frozen v2’ AI Server Chip to Improve Gemini Performance and Power Efficiency.

By Jeet Nirmal

Source: Janta Scope

Google Eyes Specialized AI Hardware for Gemini

Reports suggest that Google is designing a dedicated server chip specifically optimized for Gemini AI workloads instead of relying solely on general-purpose AI accelerators.

Unlike conventional processors that are built to support a wide variety of AI models, the proposed "Frozen v2" architecture would embed portions of Gemini's structure into the silicon itself. This approach is intended to reduce computational overhead and improve the efficiency of serving AI responses.

Addressing Growing AI Compute Demands

The reported project comes as demand for AI computing infrastructure continues to surge.

According to the report, Google has faced increasing pressure on its internal AI computing resources, with capacity constraints reportedly affecting parts of its cloud business. By creating hardware tailored specifically for Gemini, the company aims to process more AI requests while reducing power consumption and infrastructure costs.

Engineers reportedly estimate that the new chip could deliver substantially higher token-processing efficiency than Google's current generation of custom AI processors, although the design remains under development.

A New Family of AI Chips

Rather than replacing Google's TPU lineup, Frozen v2 is expected to become a separate class of specialized AI processors.

While TPUs are designed to support a broad range of machine learning workloads, the new chip would focus primarily on Gemini inference. This specialization could allow Google to reduce unnecessary processing steps and move data more efficiently within the hardware.

Reports indicate that the earliest deployment could take place around 2028, though the project remains in its early stages and design decisions are still being finalized.

Background

Google has spent years developing its own AI hardware through its Tensor Processing Unit (TPU) program, allowing the company to reduce reliance on third-party chips for many of its machine learning workloads.

As AI models become larger and more computationally demanding, major technology companies are increasingly investing in custom silicon tailored for specific applications. Specialized hardware is viewed as one of the most effective ways to lower operating costs, improve response times and expand AI services at scale.

Why This Development Matters

The reported Frozen v2 project reflects a broader shift across the AI industry toward designing hardware alongside software rather than treating them as separate technologies.

If successful, the chip could help Google serve Gemini models more efficiently, reduce energy consumption in data centers and strengthen its competitiveness in AI infrastructure. It could also lessen dependence on third-party AI processors for certain workloads while improving the scalability of Google's cloud AI services.

Balanced Analysis

Google has not officially confirmed the details of the reported project, and the chip remains under development.

If the reported plans move forward, Frozen v2 could represent an important step toward highly specialized AI hardware designed for a single model family. Such an approach may deliver significant efficiency gains but could also reduce flexibility if future AI architectures evolve in unexpected ways.

For now, the report highlights how leading AI companies are increasingly competing not only through better models but also through custom-built computing infrastructure capable of supporting those models more efficiently.

Related

More stories

Amazon Explores $8 Billion Financing Structure for Nvidia AI Chips

Amazon is reportedly exploring an unusual financing arrangement involving about $8 billion worth of Nvidia's advanced Grace Blackwell AI chips. The proposed structure would transfer thousands of chips into a special-purpose vehicle backed by outside investors, with Amazon continuing to use the processors through a lease arrangement.

AI NEWS

Amazon Explores $8 Billion Financing Structure for Nvidia AI Chips

ChatGPT Adds AI-Powered Virtual Try-On, Letting Shoppers Preview Clothes on Themselves

OpenAI has expanded ChatGPT’s shopping capabilities with an AI-powered virtual try-on feature that generates previews of users wearing clothing and accessories. Users can upload a selfie, try items surfaced in ChatGPT shopping results, or provide their own product image, while a new Favorites feature allows products to be saved for later.

AI NEWS

ChatGPT Adds AI-Powered Virtual Try-On, Letting Shoppers Preview Clothes on Themselves

Google Unveils Gemini 4 Argon, New Frontier AI Model Built for Complex Work and Cyber Defense

Google has introduced Gemini 4 Argon, its new frontier artificial-intelligence model designed for long, complex workflows across software engineering, finance, legal work and cybersecurity. The model is initially being made available to a limited group of trusted cyber defenders rather than the general public, as Google takes a phased approach to deployment and safety testing.

AI NEWS

Google Unveils Gemini 4 Argon, New Frontier AI Model Built for Complex Work and Cyber Defense

Broadcom Could Lend Anthropic Up to $42 Billion as AI Infrastructure Spending Accelerates

Broadcom has agreed to provide Anthropic with access to as much as $42 billion in financing for infrastructure spending, according to details disclosed in Anthropic's IPO documents. The arrangement deepens an already significant relationship between the semiconductor company and the AI developer as demand for computing capacity continues to rise.

AI NEWS

Broadcom Could Lend Anthropic Up to $42 Billion as AI Infrastructure Spending Accelerates

US Lawmaker Presses Major AI Companies Over Possible Chinese Access to Model Weights

U.S. Representative Ro Khanna has asked several leading American artificial intelligence companies to disclose known attempts by China or other hostile actors to gain unauthorized access to their AI model weights, putting cybersecurity around frontier AI systems under renewed congressional scrutiny.

AI NEWS

US Lawmaker Presses Major AI Companies Over Possible Chinese Access to Model Weights

IndiaAI Mission May Be Recalibrated as GPU Supply and Rising Costs Test Compute Expansion

The government is reportedly considering changes to the IndiaAI Mission's compute strategy after delays in GPU availability and rising hardware costs created a gap between committed and currently accessible capacity. The development highlights the challenges India faces as it tries to build affordable AI infrastructure while remaining dependent on global suppliers for advanced processors.

AI NEWS

IndiaAI Mission May Be Recalibrated as GPU Supply and Rising Costs Test Compute Expansion