Inception: Mercury 2.5

Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...

Inception: Mercury 2.5 — Inception. Context window 260,000 tokens. API pricing (per 1M tokens): Input $0.04, Output $0.15. License: Proprietary.

Benchmarks · API pricing (per 1M tokens)

ProviderInception
API model IDinception/mercury-2.5
TierFrontier
Context window260,000 tokens
Max output65,536 tokens
LicenseProprietary
API pricing (per 1M tokens) · Input$0.04
API pricing (per 1M tokens) · Output$0.15
Compare in the full catalog