Newest Gemini Flash for fast, cheaper generation with thinking tokens billed as output. Use for high-volume Google workloads while the intro rate lasts, then recheck the Jan 2027 step-up.
Privacy first counting
TokenCalculator counts tokens in your browser when an Exact encoding is available. Prompt text does not need to leave your device for local counting.
Pricing snapshot
USD per 1 million tokens. Verify against the provider source before you lock budget.
| Rate | USD |
|---|---|
| Input / 1M | $0.750 |
| Output / 1M | $3.75 |
| Cached input / 1M | $0.075 |
| Batch input / 1M | $0.375 |
| Batch output / 1M | $1.88 |
Model details
| Detail | Value |
|---|---|
| Provider | |
| Context window | 1.0M |
| Tokenizer | Approx (Gemini family) |
| Last verified | 2026-09-18 |
| API model id | gemini-3.8-flash |
Notes
Introductory paid rates through 2026-12-31. From 2027-01-01 standard rises to $1.50 input / $7.50 output (cache $0.15). Output price includes thinking tokens.
More Google models
Same provider catalog. Use these when you are comparing rates within one stack.
- Gemini 1.0 Pro: $0.500 in / $1.50 out
- Gemini 3 Flash: $0.500 in / $3.00 out
- Gemini 2.5 Flash: $0.300 in / $2.50 out
- Gemini 1.5 Pro: $1.25 in / $5.00 out
- Gemini 2.0 Pro Exp: $1.25 in / $5.00 out
- Gemini 2.5 Pro: $1.25 in / $10.00 out
- Gemini 3.1 Flash-Lite: $0.250 in / $1.50 out
- Gemma 3 27B: $0.200 in / $0.400 out
Featured comparisons
Browse the full set on Compare.
Related tools
- Estimate API cost
- Count tokens
- Check context fit
- Estimate cache savings
- Compare batch pricing
- Compare tokenizers