Mercury Coder Small

Model by Inception AI. A diffusion large language model that runs incredibly quickly (500+ tokens/second) while matching Claude 3.5 Haiku and GPT-4o-mini. 1st in speed on Copilot arena, and matching 2nd in quality.

Pricing

Auto routing · per 1M tokens
Input
$0.25
Output
$1.00
Compare provider prices

Specifications

Context window
32.8K
Max output
16.4K

Benchmarks

No public benchmark scores for this model yet.

Providers

Auto routing is available for this model. Explicit provider selection is not available.

Loading provider options…