Mercury 2.5 from @_inception_ai is now generally available on OpenRouter.
It's the fastest model on our Fastest Models leaderboard that doesn't require specialized hardware.
A diffusion LLM: it decodes tokens in parallel rather than one at a time. The gain is largest on code-heavy prompts.
Use it here: https://openrouter.ai/inception/mercury-2.5