mercury

Mercury Coder Small

Model by Inception AI. A diffusion large language model that runs incredibly quickly (500+ tokens/second) while matching Claude 3.5 Haiku and GPT-4o-mini. 1st in speed on Copilot arena, and matching 2nd in quality.

Fresh · fetched 2026-08-17T02:11:57.739Z · Models.dev(opens in a new tab)

Minimum context32K
Output16K
Providers2
CapabilitiesTool calling

Provider offers

ProvidersInput / 1MOutput / 1MCache read / 1MContext / outputUpdated
NanoGPTmercury-coder-small$0.25$1.00$0.1332K2024-01-01
Vercel AI Gatewayinception/mercury-coder-small$0.25$1.00Unavailable32K2025-02-26

Lowest reported provider price in USD per 1M tokens; unavailable is not free. Provider documentation remains authoritative. Catalogue presence is not an uptime, security, or quality endorsement.

Source metadata and limitations

Capabilities are source-reported. Tool calling does not prove quality in a particular harness. Prices and limits can change.

Modalities: texttext