AI Agent News
Track major events, funding, model releases and breakthroughs across the AI Agent landscape
Latest industry news
Track major events, funding, model releases and breakthroughs across the AI Agent landscape
Key events timeline
OpenClaw erupts on GitHub
OpenClaw hits global GitHub Top 10 in 10 days, outpacing the Linux kernel star growth
Meta acquires Manus for $2B
Meta acquires Manus AI for $2B, locking in the general-purpose Agent race
DeepSeek-V3 open-sourced
The value king, at just 5% of GPT-4 cost
Manus goes viral overnight
The world's first general-purpose AI Agent draws unprecedented attention
OpenAI Deep Research
OpenAI ships a deep-research Agent that generates professional reports in one click
MCP Servers pass 500
The MCP ecosystem erupts — 500+ servers built in 3 months
DeepSeek-R1 stuns the world
Open-source reasoning model at just 3% of OpenAI cost, reshaping the global AI landscape
MCP protocol born
Anthropic releases the Model Context Protocol, the de facto standard for Agent interfaces
Claude Computer Use
Anthropic lets AI directly control the computer screen for the first time, opening a new paradigm
Replit Agent full-stack automation
Natural language to a shipped product, aimed at non-engineers
Cursor ARR passes $100M
The fastest-growing SaaS ever, the new king of AI coding tools
Claude 3.5 tops SWE-bench
The strongest coding AI, bug-fixing at a junior engineer level
Devin launches
The world's first autonomous AI software engineer, able to complete full coding tasks on its own
AI Inference Chip Demand Creates New Semiconductor Supply Crisis
Surging demand for AI inference chips from cloud providers and edge device manufacturers creates supply constraints. TSMC and Samsung announce $50B+ expansion programs to meet projected 300% demand growth.
NVIDIA Announces Blackwell Ultra GPU: 20x Performance Improvement for LLM Inference
NVIDIA unveils Blackwell Ultra architecture promising 20x throughput improvement for LLM inference workloads, with first deliveries to cloud providers expected in Q3 2025.
NVIDIA Blackwell B200 Ships in Volume: 2.5x H100 Performance Transforms AI Training Economics
NVIDIA's Blackwell B200 GPU has begun shipping in volume to major cloud providers and hyperscalers, delivering 2.5x the training throughput of the H100 for transformer models. The B200 features 192GB HBM3e memory and 10 PB/s NVLink bandwidth in an 8-GPU NVL72 configuration. AWS, Google Cloud, Microsoft Azure, and Oracle Cloud have announced Blackwell-based instances. First customer benchmarks show GPT-4 class model training 40% faster than H100 DGX at comparable cost, significantly changing the economics of large-scale AI development.
AMD Instinct MI325X Ships: 288GB HBM3e Challenges NVIDIA H200 for LLM Inference
AMD has begun shipping the Instinct MI325X accelerator, featuring 288GB HBM3e memory—36% more than NVIDIA H200's 141GB. The additional memory allows serving larger LLM batches without quantization, potentially improving inference quality. AMD's ROCm 6.2 software stack now supports all major ML frameworks (PyTorch, JAX, TensorFlow) with competitive performance to CUDA. Microsoft Azure has deployed MI325X clusters for Azure AI services, and Oracle Cloud has announced MI325X availability. AMD claims MI325X achieves 92% of H200 performance at 80% of the cost for inference workloads.