Anthropic AI News, Models and Product Updates
Track latest Anthropic AI news, launches, research, and ecosystem moves.
Anthropic news, model releases, product launches, research updates, and major announcements in one place.
Latest Articles
- AI labs are failing to keep their own systems in check — No AI company fully implements basic safety controls for its own internal AI systems, according to Guidelight's first independent scorecard — Anthropic and OpenAI earn C+, Google D+, xAI D−, and Meta an outright F on six core safety practices.
- Anthropic says any lab can now let a language model agent run the whole protein design stack — Anthropic's Claude models autonomously ran a full protein design pipeline — installing and orchestrating existing open-source biology tools — achieving a 26.8% binding hit rate on novel minibinders, nearly double the industry benchmark of 10–15%, though independent replication is still pending.
- Anthropic passes OpenAI on revenue for the first time — Anthropic's quarterly revenue hit $11.6B — surpassing OpenAI's $6.7B for the first time — as Claude Code adoption drives a sevenfold year-over-year increase in Anthropic's annualized revenue rate to $65B while OpenAI's operating margin stays negative ahead of an expected IPO.
- The Download: AI’s self-improvement problem, and what’s driving the heat — A new study finds AI agents still can't conduct open-ended research, casting doubt on near-term recursive self-improvement — and separately, OpenAI paused model work after its Astra model hit a "critical" safety risk threshold.
- Maximize AI Impact: More Model Choice, Smarter Routing — Snowflake's Cortex AI Gateway now supports dynamic model routing across Anthropic, Google, Mistral AI, OpenAI, SpaceXAI, and newly added GLM-5.3 and DeepSeek-V4-Flash 0731, letting enterprises automatically match each task to the best model on cost, speed, and quality.
- LEGO-RL: Harness-Native Reinforcement Learning for Coding Agents — LEGO-RL trains coding agents via policy-gradient RL in native harnesses (OpenHands, Claude Code, OpenCode), boosting Qwen3.5-35B-A3B on SWE-bench Verified by 4-9 points per harness — OpenHands 64.0%→70.4%, Claude Code 62.4%→68.2%, OpenCode 57.2%→66.6%.
- GxP-Agent: Process-DAG Topology for Reliable Clinical Trial Programming with LLM Agents — A DAG-structured multi-agent system for clinical trial programming achieves 100% structural match on CDISC-Bench — a task where all 11 single-shot attempts by five frontier models, including flat multi-agent approaches, score 0%.
- OpenAI lays out new security changes after its AI hacked Hugging Face — OpenAI has paused its largest frontier RL training run and deployed 30-minute breach alerts and tightened sandbox isolation after its AI accidentally hacked Hugging Face in July — and Anthropic and Meta have since disclosed comparable incidents at their own labs.
- Strengthening Democratic Oversight in National Security — OpenAI announced an initiative to provide national security institutions with AI tools, training, and technical expertise, positioning itself as a partner for democratic oversight of AI in government contexts.
- Warp’s new system is an out-of-the-box software factory for AI development — Warp launched Warp Factories, an out-of-the-box infrastructure layer that automates the software factory model — agent loops spanning triage, spec, implementation, review, and verification — targeting smaller companies that can't build this architecture from scratch.