Direct API pricing comparison
Gemini 3.5 Flash-Lite vs Gemini 3.1 Pro Preview
Compare current direct API token prices and estimated monthly workload costs for Gemini 3.5 Flash-Lite and Gemini 3.1 Pro Preview.
2direct models
Pricing only.Gemini 3.5 Flash-Lite is cheaper in 4 of 4 comparable standard workload presets below. This is a pricing comparison only; it does not rank model quality.
| Model | Input | Cached | Output | Context | Max output | Source |
|---|---|---|---|---|---|---|
Gemini 3.5 Flash-LiteGoogle · Direct API | $0.3 | $0.03 | $2.5 | 1.05M | 66K | Direct ↗ |
Gemini 3.1 Pro PreviewGoogle · Direct API | $2 | $0.2 | $12 | 1.05M | 66K | Direct ↗ |
Same workload, different API bill
These four presets use 30 active days, standard direct rates, no retries, no reasoning-token add-on, no cache writes, and the cached-input percentage shown in each row.
| Standard workload | Gemini 3.5 Flash-Lite | Gemini 3.1 Pro Preview | Lower monthly estimate |
|---|---|---|---|
| Chat10,000/day · 2,000 in · 500 out · 25% cached | $514.5 | $2,730 | Gemini 3.5 Flash-Lite by $2,215.5 / mo |
| Coding2,000/day · 12,000 in · 1,800 out · 55% cached | $379.08 | $2,023.2 | Gemini 3.5 Flash-Lite by $1,644.12 / mo |
| Agent1,200/day · 32,000 in · 2,500 out · 65% cached | $368.42 | $2,036.16 | Gemini 3.5 Flash-Lite by $1,667.74 / mo |
| Documents400/day · 180,000 in · 2,000 out · 15% cached | $620.52 | $4,024.8 | Gemini 3.5 Flash-Lite by $3,404.28 / mo |
The main calculator includes batch mode, long-context rules, retries, reasoning tokens, cache writes, and other model-specific pricing logic that a static comparison cannot fully capture.