qwen

Qwen: QvQ Max

Compact GPT model for low-latency assistance and high-volume workloads

Fresh · fetched 2026-08-17T02:11:57.739Z · Models.dev(opens in a new tab)

Minimum context128K
Output8K
Providers3
CapabilitiesTool calling, Reasoning, Image input

Provider offers

ProvidersInput / 1MOutput / 1MCache read / 1MContext / outputUpdated
NanoGPTqvq-max$1.20$4.80$0.60128K2025-03-28
Alibaba (China)qvq-max$1.15$4.59Unavailable128K2025-03-25
Alibabaqvq-max$1.20$4.80Unavailable128K2025-03-25

Lowest reported provider price in USD per 1M tokens; unavailable is not free. Provider documentation remains authoritative. Catalogue presence is not an uptime, security, or quality endorsement.

Source metadata and limitations

Capabilities are source-reported. Tool calling does not prove quality in a particular harness. Prices and limits can change.

Modalities: text, imagetext