Labs
Every move from OpenAI, Anthropic, Google DeepMind, Meta, Mistral, xAI, DeepSeek, Qwen, Moonshot and the global frontier labs: model launches, research drops, product news.
Every move from OpenAI, Anthropic, Google DeepMind, Meta, Mistral, xAI, DeepSeek, Qwen, Moonshot and the global frontier labs: model launches, research drops, product news.
Latest in Labs
30 stories
Microsoft's SkillOpt boosts GPT-5.5 by using nothing but a trained Markdown file
Microsoft and three Chinese universities have developed SkillOpt, a method that optimizes instruction documents for AI agents using principles from traditional …

Google Research's Gemini-SQL2 tops text-to-SQL benchmarks by a wide margin
Google Research's Gemini-SQL2 turns natural language into executable SQL queries. Built on Gemini 3.1 Pro, it tops the BIRD benchmark at 80.04 percent accuracy,…

Claude Fable 5 outpaces GPT-5.5 by 13 points on FrontierMath's toughest problems
Anthropic's Claude Fable 5 hits 88 percent accuracy on the hardest FrontierMath tier, a massive jump from Opus 4.5, which sat below 10 percent in early 2026. Op…

Moonshot's open model Kimi K2.7 Code undercuts GPT-5.5 and Claude by up to 12x on price per token
Moonshot AI has released Kimi K2.7 Code, an open-weights model with one trillion parameters built for programming. It still trails GPT-5.5 and Claude Opus 4.8 i…

US government forces Anthropic to disable Claude Fable 5 and Mythos 5 for all customers worldwide
The US government has ordered Anthropic to shut down global access to Fable 5 and Mythos 5, citing alleged jailbreak risks. Anthropic is complying but pushing b…
Claude Code v2.1.176

Over half of Americans fear losing both their jobs and their independent thinking to AI, survey finds
Anthropic surveyed nearly 52,000 Americans about their hopes and fears around AI. Sixty-four percent fear job losses, and 56 percent worry about losing the abil…

OpenAI kicks off the AI price wars with flexible rate-limit resets for its Codex coding agent
OpenAI now lets Codex users bank their rate-limit resets and trigger them manually instead of watching them expire on a fixed schedule. If you hit your usage ca…

Anthropic's Claude Fable 5 costs twice as much for 5.7 percent more performance
Claude Fable 5 tops the Artificial Analysis Intelligence Index with 64.9 points and sets records in five of ten benchmarks. But the gain over Opus 4.8 is just 5…

Mistral AI seeks 3 billion euros to fund its European AI push
French AI startup Mistral AI is negotiating a new funding round of around 3 billion euros at a valuation of approximately 20 billion euros. The article Mistral …

Google files first joint lawsuit with FBI over Chinese AI scam network, OpenAI blocks PRC influence clusters
Within days of each other, Google and OpenAI separately exposed operations allegedly originating in China that use AI for fraud and covert influence campaigns. …

The AI industry's platform trap is starting to look a lot like Microsoft's
Anthropic is throttling its new Mythos model for certain tasks while building apps that directly compete with its largest customers. Customers, partners, and in…
OpenAI Agents SDK v0.15.2
OpenAI Agents SDK v0.15.3
OpenAI Agents SDK v0.16.0
OpenAI Agents SDK v0.16.1
OpenAI Agents SDK v0.17.0
OpenAI Agents SDK v0.17.1
OpenAI Agents SDK v0.17.2
OpenAI Agents SDK v0.17.3
OpenAI Agents SDK v0.17.4
Claude Code v2.1.166
Claude Code v2.1.169
Claude Code v2.1.170
Claude Code v2.1.172
OpenAI Agents SDK v0.17.5
Claude Code v2.1.173
Claude Code v2.1.174
Claude Code v2.1.175

Anthropic built a model too risky to release
and Meta makes an unexpected entry…