qwen
Qwen3.8 2.4T A95B
Open-weight sparse MoE (2.4T total, 95B active), the open-weight twin of Qwen3.8 Max for coding, research, complex reasoning, and agentic workflows
Fresh · fetched 2026-08-17T02:11:57.739Z · Models.dev(opens in a new tab)
Minimum context256K
Output32K
Providers1
CapabilitiesTool calling, Reasoning
Provider offers
| Providers | Input / 1M | Output / 1M | Cache read / 1M | Context / output | Updated |
|---|---|---|---|---|---|
Inferact/Qwen3.8-2.4T-A95B-NVFP4 | $2.00 | $6.00 | $0.20 | 256K | 2026-08-12 |
Lowest reported provider price in USD per 1M tokens; unavailable is not free. Provider documentation remains authoritative. Catalogue presence is not an uptime, security, or quality endorsement.
Source metadata and limitations
Capabilities are source-reported. Tool calling does not prove quality in a particular harness. Prices and limits can change.
Modalities: text → text