Cheaper GPT-4.1 variant for the same long-context shape. Use for RAG and assistants where Mini quality is enough and you want Exact OpenAI token counts.
Privacy first counting
TokenCALC counts tokens in your browser when an Exact encoding is available. Prompt text does not need to leave your device for local counting.
Pricing snapshot
USD per 1 million tokens. Verify against the provider source before you lock budget.
| Rate | USD |
|---|---|
| Input / 1M | $0.400 |
| Output / 1M | $1.60 |
| Cached input / 1M | $0.100 |
| Batch input / 1M | $0.200 |
| Batch output / 1M | $0.800 |
Model details
| Detail | Value |
|---|---|
| Provider | OpenAI |
| Context window | 1M |
| Tokenizer | Exact (o200k_base) |
| Last verified | 2026-08-02 |
| API model id | gpt-4.1-mini |