llama

Llama-4-Maverick-17B-128E-Instruct-FP8

Open multimodal Llama model for strong reasoning and fast responses

Fresh · fetched 2026-08-17T02:11:57.739Z · Models.dev(opens in a new tab)

Minimum context1.05M
Output43K
Providers9
CapabilitiesTool calling, Image input

Provider offers

ProvidersInput / 1MOutput / 1MCache read / 1MContext / outputUpdated
Deep Inframeta-llama/Llama-4-Maverick-17B-128E-Instruct-FP8$0.20$0.80Unavailable1.05M2025-04-05
Llamallama-4-maverick-17b-128e-instruct-fp8$0.00$0.00Unavailable128K2025-04-05
IO.NETmeta-llama/Llama-4-Maverick-17B-128E-Instruct-FP8$0.15$0.60$0.075430K2025-01-15
Azure Cognitive Servicesllama-4-maverick-17b-128e-instruct-fp8$0.25$1.00Unavailable1M2025-04-05
Charm Hyperllama-4-maverick-17b-128e-instruct-fp8$0.27$0.90Unavailable430K2026-07-22
Azurellama-4-maverick-17b-128e-instruct-fp8$0.25$1.00Unavailable1M2025-04-05
watsonx.aimeta-llama/llama-4-maverick-17b-128e-instruct-fp8$0.37$1.48Unavailable128K2025-04-05
NovitaAImeta-llama/llama-4-maverick-17b-128e-instruct-fp8$0.27$0.85Unavailable1.05M2025-04-06
Abacusmeta-llama/Llama-4-Maverick-17B-128E-Instruct-FP8$0.14$0.59Unavailable1.05M2025-04-05

Lowest reported provider price in USD per 1M tokens; unavailable is not free. Provider documentation remains authoritative. Catalogue presence is not an uptime, security, or quality endorsement.

Source metadata and limitations

Capabilities are source-reported. Tool calling does not prove quality in a particular harness. Prices and limits can change.

Modalities: text, imagetext