gemini-flash-lite

Gemini Flash-Lite Latest

Low-latency Gemini model for high-volume multimodal and agent workloads

Fresh · fetched 2026-08-17T02:11:57.739Z · Models.dev(opens in a new tab)

Minimum context1.05M
Output64K
Providers5
CapabilitiesTool calling, Reasoning, Image input

Provider offers

ProvidersInput / 1MOutput / 1MCache read / 1MContext / outputUpdated
OrcaRoutergoogle/gemini-flash-lite-latest$0.25$1.50$0.0251.05M2026-05-07
NanoGPTgoogle/gemini-flash-lite-latest$0.30$2.50$0.0301.05M2026-05-07
Vertexgemini-flash-lite-latest$0.25$1.50$0.0251.05M2026-05-07
Googlegemini-flash-lite-latest$0.25$1.50$0.0251.05M2026-05-07
Merge Gatewaygoogle/gemini-flash-lite-latest$0.25$1.50$0.0251.05M2026-05-07

Lowest reported provider price in USD per 1M tokens; unavailable is not free. Provider documentation remains authoritative. Catalogue presence is not an uptime, security, or quality endorsement.

Source metadata and limitations

Capabilities are source-reported. Tool calling does not prove quality in a particular harness. Prices and limits can change.

Modalities: text, image, video, audio, pdftext