LLMPrice
Updated daily
Direct API pricing comparison

Gemini 3.7 Flash vs Gemini 3.5 Flash-Lite

Compare current direct API token prices and estimated monthly workload costs for Gemini 3.7 Flash and Gemini 3.5 Flash-Lite.

2direct models
Pricing only.Gemini 3.5 Flash-Lite is cheaper in 4 of 4 comparable standard workload presets below. This is a pricing comparison only; it does not rank model quality.
Published direct rates

Token pricing side by side

USD per 1M text tokens
ModelInputCachedOutputContextMax outputSource
Gemini 3.7 FlashGoogle · Direct API
$0.75$0.075$3.751.05M66KDirect
Gemini 3.5 Flash-LiteGoogle · Direct API
$0.3$0.03$2.51.05M66KDirect
Monthly estimates

Same workload, different API bill

These four presets use 30 active days, standard direct rates, no retries, no reasoning-token add-on, no cache writes, and the cached-input percentage shown in each row.

Standard workloadGemini 3.7 FlashGemini 3.5 Flash-LiteLower monthly estimate
Chat10,000/day · 2,000 in · 500 out · 25% cached$911.25$514.5Gemini 3.5 Flash-Lite by $396.75 / mo
Coding2,000/day · 12,000 in · 1,800 out · 55% cached$677.7$379.08Gemini 3.5 Flash-Lite by $298.62 / mo
Agent1,200/day · 32,000 in · 2,500 out · 65% cached$696.06$368.42Gemini 3.5 Flash-Lite by $327.64 / mo
Documents400/day · 180,000 in · 2,000 out · 15% cached$1,491.3$620.52Gemini 3.5 Flash-Lite by $870.78 / mo
Custom workload

Use your own traffic

The main calculator includes batch mode, long-context rules, retries, reasoning tokens, cache writes, and other model-specific pricing logic that a static comparison cannot fully capture.

Related comparisons

Compare nearby alternatives