LLMPrice
Updated daily
Direct API pricing comparison

Gemini 3.7 Flash vs Gemini 3.1 Pro Preview

Compare current direct API token prices and estimated monthly workload costs for Gemini 3.7 Flash and Gemini 3.1 Pro Preview.

2direct models
Pricing only.Gemini 3.7 Flash is cheaper in 4 of 4 comparable standard workload presets below. This is a pricing comparison only; it does not rank model quality.
Published direct rates

Token pricing side by side

USD per 1M text tokens
ModelInputCachedOutputContextMax outputSource
Gemini 3.7 FlashGoogle · Direct API
$0.75$0.075$3.751.05M66KDirect
Gemini 3.1 Pro PreviewGoogle · Direct API
$2$0.2$121.05M66KDirect
Monthly estimates

Same workload, different API bill

These four presets use 30 active days, standard direct rates, no retries, no reasoning-token add-on, no cache writes, and the cached-input percentage shown in each row.

Standard workloadGemini 3.7 FlashGemini 3.1 Pro PreviewLower monthly estimate
Chat10,000/day · 2,000 in · 500 out · 25% cached$911.25$2,730Gemini 3.7 Flash by $1,818.75 / mo
Coding2,000/day · 12,000 in · 1,800 out · 55% cached$677.7$2,023.2Gemini 3.7 Flash by $1,345.5 / mo
Agent1,200/day · 32,000 in · 2,500 out · 65% cached$696.06$2,036.16Gemini 3.7 Flash by $1,340.1 / mo
Documents400/day · 180,000 in · 2,000 out · 15% cached$1,491.3$4,024.8Gemini 3.7 Flash by $2,533.5 / mo
Custom workload

Use your own traffic

The main calculator includes batch mode, long-context rules, retries, reasoning tokens, cache writes, and other model-specific pricing logic that a static comparison cannot fully capture.

Related comparisons

Compare nearby alternatives