GPT-4.1 Nano vs Gemini 2.5 Flash-Lite

Side by side pricing and context for GPT-4.1 Nano and Gemini 2.5 Flash-Lite. Rates are curated catalog values with last verified dates on each model page.

GPT-4.1 Nano

Budget GPT-4.1 tier for classification, routing, and high-QPS extraction. Estimate monthly volume early: small per-token prices still add up at scale.

Input $0.100 / Output $0.400 per 1M

Context 1M · Exact (o200k_base)

Gemini 2.5 Flash-Lite

Cheapest Gemini Flash tier for routing, moderation, and simple extraction. Use when Flash quality is more than you need.

Input $0.100 / Output $0.400 per 1M

Context 1M · Approx (Gemini family)


Quick take

  • Input rates match on this snapshot.
  • Output rates match on this snapshot.
  • Both last verified on their model pages (2026-08-02 / 2026-08-02).

Rate table

MetricGPT-4.1 NanoGemini 2.5 Flash-Lite
Input / 1M$0.100$0.100
Output / 1M$0.400$0.400
Cached input / 1M$0.025$0.010
Context1M1M
TokenizerExact (o200k_base)Approx (Gemini family)