mercury
Mercury 2
Compact GPT model for low-latency assistance and high-volume workloads
Fresh · fetched 2026-08-17T02:11:57.739Z · Models.dev(opens in a new tab)
Minimum context128K
Output128K
Providers6
CapabilitiesTool calling, Reasoning
Provider offers
| Providers | Input / 1M | Output / 1M | Cache read / 1M | Context / output | Updated |
|---|---|---|---|---|---|
mercury-2 | $0.25 | $0.75 | $0.025 | 128K | 2024-01-01 |
mercury-2 | $0.25 | $0.75 | $0.025 | 128K | 2026-02-24 |
inception/mercury-2 | $0.25 | $0.75 | $0.025 | 128K | 2026-03-04 |
inception/mercury-2 | $0.25 | $0.75 | $0.025 | 128K | 2026-03-06 |
inception/mercury-2 | $0.25 | $0.75 | $0.025 | 128K | 2026-03-04 |
mercury-2 | $0.31 | $0.94 | $0.031 | 128K | 2026-06-11 |
Lowest reported provider price in USD per 1M tokens; unavailable is not free. Provider documentation remains authoritative. Catalogue presence is not an uptime, security, or quality endorsement.
Source metadata and limitations
Capabilities are source-reported. Tool calling does not prove quality in a particular harness. Prices and limits can change.
Modalities: text → text