Side by side pricing and context for GPT-4.1 Nano and Gemini 2.5 Flash-Lite. Rates are curated catalog values with last verified dates on each model page.
GPT-4.1 Nano
Budget GPT-4.1 tier for classification, routing, and high-QPS extraction. Estimate monthly volume early: small per-token prices still add up at scale.
Input $0.100 / Output $0.400 per 1M
Context 1M · Exact (o200k_base)
Gemini 2.5 Flash-Lite
Cheapest Gemini Flash tier for routing, moderation, and simple extraction. Use when Flash quality is more than you need.
Input $0.100 / Output $0.400 per 1M
Context 1M · Approx (Gemini family)
Quick take
- Input rates match on this snapshot.
- Output rates match on this snapshot.
- Both last verified on their model pages (2026-08-02 / 2026-08-02).
Rate table
| Metric | GPT-4.1 Nano | Gemini 2.5 Flash-Lite |
|---|---|---|
| Input / 1M | $0.100 | $0.100 |
| Output / 1M | $0.400 | $0.400 |
| Cached input / 1M | $0.025 | $0.010 |
| Context | 1M | 1M |
| Tokenizer | Exact (o200k_base) | Approx (Gemini family) |