AI Agent News
Track major events, funding, model releases and breakthroughs across the AI Agent landscape
Latest industry news
Track major events, funding, model releases and breakthroughs across the AI Agent landscape
Key events timeline
OpenClaw erupts on GitHub
OpenClaw hits global GitHub Top 10 in 10 days, outpacing the Linux kernel star growth
Meta acquires Manus for $2B
Meta acquires Manus AI for $2B, locking in the general-purpose Agent race
DeepSeek-V3 open-sourced
The value king, at just 5% of GPT-4 cost
Manus goes viral overnight
The world's first general-purpose AI Agent draws unprecedented attention
OpenAI Deep Research
OpenAI ships a deep-research Agent that generates professional reports in one click
MCP Servers pass 500
The MCP ecosystem erupts — 500+ servers built in 3 months
DeepSeek-R1 stuns the world
Open-source reasoning model at just 3% of OpenAI cost, reshaping the global AI landscape
MCP protocol born
Anthropic releases the Model Context Protocol, the de facto standard for Agent interfaces
Claude Computer Use
Anthropic lets AI directly control the computer screen for the first time, opening a new paradigm
Replit Agent full-stack automation
Natural language to a shipped product, aimed at non-engineers
Cursor ARR passes $100M
The fastest-growing SaaS ever, the new king of AI coding tools
Claude 3.5 tops SWE-bench
The strongest coding AI, bug-fixing at a junior engineer level
Devin launches
The world's first autonomous AI software engineer, able to complete full coding tasks on its own
Nature and Science Retract 150 AI-Generated Papers With Fabricated Data
Major scientific journals retract 150 papers found to contain AI-generated fabricated data and images. Publishers announce mandatory AI-assisted authorship disclosure and enhanced fraud detection screening.
AI Synthetic Media Detection in 2025 Elections: Global Challenges and Solutions
Election security agencies report detecting over 100,000 AI-generated synthetic media pieces targeting elections across 15 countries. Coalitions form to deploy detection tools and voter education campaigns.
Tech Giants Launch Coalition for AI-Generated Content Detection Standards
Microsoft, Google, Meta, and Adobe launch the Content Authenticity Initiative for AI, establishing watermarking and provenance standards for AI-generated images, video, and audio.
Anthropic Publishes Updated Model Spec: New Guidelines for AI Behavior
Anthropic releases comprehensive update to Claude Model Spec, detailing new guidelines for handling sensitive topics, improved calibration for confidence expressions, and enhanced corrigibility principles.
Anthropic's Mechanistic Interpretability Research Finds 'Features' in Claude's Reasoning
Anthropic published landmark interpretability research identifying thousands of 'features' — linear representations of concepts — inside Claude's neural network activations. Researchers found features corresponding to concepts like 'the Golden Gate Bridge,' 'code bugs,' and emotional states. More concerning: researchers identified features active during deceptive responses. This work brings the field closer to explaining why LLMs behave as they do, a necessary precondition for reliable AI safety guarantees.
Anthropic Achieves Breakthrough in Mechanistic Interpretability Research
Anthropic researchers publish landmark paper on mechanistic interpretability, successfully mapping how Claude represents concepts internally and identifying circuits responsible for safety behaviors.
OpenAI Establishes Safety and Security Committee, Releases Enhanced Model Security Guidelines
OpenAI has established a permanent Safety and Security Committee following months of internal and external pressure around AI safety practices. The committee has published new Model Security Guidelines covering red teaming requirements, catastrophic risk thresholds, and mandatory security reviews before major model releases. OpenAI also announced a formal vulnerability disclosure program for AI-specific security issues, with bounties up to $100,000 for critical AI safety vulnerabilities.
Anthropic Publishes Constitutional AI Safety Update: Claude 3.7 Security and Jailbreak Resistance
Anthropic has released its most comprehensive AI safety update to date, detailing Constitutional AI improvements in Claude 3.7 that reduce harmful output by 89% and jailbreak attempts by 94% compared to Claude 2. The report includes new safety benchmarks, red team findings from 200+ external researchers, and a technical specification of the Responsible Scaling Policy (RSP) thresholds that would trigger halting development of more powerful models. Anthropic also published ASL-3 requirements—the safety bar required before deploying models with potential for CBRN uplift.