Side by side pricing and context for o4-mini and GPT-5 Mini. Rates are curated catalog values with last verified dates on each model page.
o4-mini
Smaller reasoning model when o3 is too expensive. Use for math, planning, and agent steps where you need chain-of-thought style work on a budget.
Input $1.10 / Output $4.40 per 1M
Context 200K · Exact (o200k_base)
GPT-5 Mini
Lower-cost GPT-5 tier for high-volume chat and light reasoning. Good default when GPT-5 is overkill but you still want the GPT-5 family tokenizer (Exact in TokenCALC).
Input $0.250 / Output $2.00 per 1M
Context 128K · Exact (o200k_base)
Quick take
- GPT-5 Mini is cheaper on input in this catalog snapshot.
- GPT-5 Mini is cheaper on output in this catalog snapshot.
- Both last verified on their model pages (2026-08-02 / 2026-08-02).
Rate table
| Metric | o4-mini | GPT-5 Mini |
|---|---|---|
| Input / 1M | $1.10 | $0.250 |
| Output / 1M | $4.40 | $2.00 |
| Cached input / 1M | $0.275 | $0.025 |
| Context | 200K | 128K |
| Tokenizer | Exact (o200k_base) | Exact (o200k_base) |