LLM Research & News Mercury 2.5 Is Fast. That Is the Claim You Have to Test
Inception's diffusion LLM hit 1,107 tokens per second and a 260K window. The launch pricing is cheap. The architecture is the bet, not another Flash-class incremental drop.