Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception
Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Supports a context window of 260,000 tokens. Supports extended reasoning. Supports tool calling. Input priced at $0.04 per million tokens. Weights are not publicly released; access is via the provider API.
Inception
api
paid
Others in the same category, ranked by how often they are opened.