中文

AI Agent News

Track major events, funding, model releases and breakthroughs across the AI Agent landscape

AI Agent updates

Latest industry news

Track major events, funding, model releases and breakthroughs across the AI Agent landscape

Key events timeline

2026-01

OpenClaw erupts on GitHub

OpenClaw hits global GitHub Top 10 in 10 days, outpacing the Linux kernel star growth

2025-12

Meta acquires Manus for $2B

Meta acquires Manus AI for $2B, locking in the general-purpose Agent race

2025-04

DeepSeek-V3 open-sourced

The value king, at just 5% of GPT-4 cost

2025-03

Manus goes viral overnight

The world's first general-purpose AI Agent draws unprecedented attention

2025-02

OpenAI Deep Research

OpenAI ships a deep-research Agent that generates professional reports in one click

2025-02

MCP Servers pass 500

The MCP ecosystem erupts — 500+ servers built in 3 months

2025-01

DeepSeek-R1 stuns the world

Open-source reasoning model at just 3% of OpenAI cost, reshaping the global AI landscape

2024-11

MCP protocol born

Anthropic releases the Model Context Protocol, the de facto standard for Agent interfaces

2024-10

Claude Computer Use

Anthropic lets AI directly control the computer screen for the first time, opening a new paradigm

2024-09

Replit Agent full-stack automation

Natural language to a shipped product, aimed at non-engineers

2024-08

Cursor ARR passes $100M

The fastest-growing SaaS ever, the new king of AI coding tools

2024-06

Claude 3.5 tops SWE-bench

The strongest coding AI, bug-fixing at a junior engineer level

2024-03

Devin launches

The world's first autonomous AI software engineer, able to complete full coding tasks on its own

IndustryJul 21, 2026

Google Reportedly Developing Frozen v2 Chip: Hardcoding Gemini Architecture into Silicon, Energy Efficiency Could Surpass TPU by Tenfold

According to an exclusive report by The Information, Google is developing a server AI chip codenamed "Frozen v2" that aims to permanently etch parts of the underlying architecture of the Gemini model into silicon, rather than serving as a general-purpose computing platform like traditional chips. The project is led by Jeff Dean, Chief Scientist at Google DeepMind, and is an iteration of the earlier Frozen project. ## Core Design: From "Running Models" to "Growing Models" The core idea of Frozen v2 is to "harden the architecture while preserving weights." Unlike general-purpose chips such as NVIDIA GPUs or Google TPUs, Frozen v2 is custom-built solely for the Gemini model, embedding critical computation paths directly into circuits, eliminating redundant steps like real-time scheduling and data movement. Google employees estimate that, measured by tokens processed per watt, its energy efficiency could be 6 to 10 times higher than the latest generation of TPUs. - **Earlier Frozen project**: Attempted to burn model weights directly into the chip, but was shelved because the hardened weights could not adapt to model updates. - **Frozen v2 compromise**: Only the underlying computational blueprint is hardened; weights can still be updated. As long as the Gemini architecture does not undergo major changes, the chip can be used continuously. - **Engineering trade-off**: Google is still debating the degree of hardening—more locking yields higher efficiency but less flexibility. ## Compute Crisis: Why Google Bets on Specialized Chips Google faces a severe compute shortage, which has already forced Google Cloud to reject orders from external customers. In June 2024, Google signed a contract with SpaceX, paying $920 million per month to lease 110,000 NVIDIA GPUs, with the contract lasting until 2029. - **Industry trend**: The entire industry is betting on inference chips, including startups like SambaNova and d-Matrix, as well as giants like OpenAI, Microsoft (Maia), and Amazon (Trainium/Inferentia). NVIDIA also acquired Groq's technology license for $20 billion in December 2023. - **Extreme route**: Canadian startup Taalas has already pursued the "model hardwiring" approach, raising over $200 million. ## Risks and Bets: Architecture Convergence or Rigidity? The cost of Frozen v2 is loss of flexibility: only if subsequent Gemini models use the same underlying architecture can the chip continue to work. Google is betting that model architectures will not undergo fundamental changes in the coming years. - **Model progress**: Gemini 3.5 Pro (codenamed Cappuccino) has been delayed three times due to subpar coding capabilities, originally scheduled for June 2024 but now postponed by several months. - **Industry context**: A Wharton professor pointed out the "disappointment trap of next-generation giant models," where the returns from piling up data and compute are diminishing, and the pace of architectural iteration is slowing. - **Deployment timeline**: The chip is expected to be deployed as early as 2028, with production volume not reaching TPU scale, more like a limited experiment. ## Implications and Insights Frozen v2 marks the evolution of AI chips from general-purpose to extreme specialization, similar to the shift from CPU to ASIC in Bitcoin mining. However, this could hinder the development of new architectures (e.g., non-Transformer models) due to the lack of compatible hardware. Google's bet reflects the industry's expectation of architectural convergence, but if a paradigm shift occurs, Frozen v2's high efficiency could instantly become a technical liability.