Mercury 2: The First Reasoning Diffusion Language Model — How Inception Labs Is Reinventing AI Architecture
On February 24, 2026, AI startup Inception Labs launched Mercury 2 — the world's first Reasoning Diffusion Language Model (dLLM). The benchmark numbers are staggering: 1,009 tokens per second, 5x faster than Claude 4.5 Haiku, at pricing that undercuts competitors by 2–4x. This isn't a faster version of ChatGPT or Claude. It's a fundamentally different approach to how AI generates text.
Founded by researchers from Stanford, UCLA, and Cornell who contributed to foundational diffusion model research, Inception Labs has spent years applying diffusion — the technique behind image generators like Stable Diffusion — to language. Mercury 2 is the production-ready result.
Premium Content
You've read all your free articles today. Subscribe to continue reading.
You've used 3 of 3 free articles today.
Subscribe NowAlready subscribed? Sign in




Comments (0)
Be the first to comment!