Side by side pricing and context for Llama 3.2 1B (Groq) and Llama 3.2 3B (Groq). Rates are curated catalog values with last verified dates on each model page.
Llama 3.2 1B (Groq)
Smallest Llama on Groq for extreme high-QPS routing. Use only when quality bars are low and latency/cost dominate.
Input $0.040 / Output $0.040 per 1M
Context 128K · Approx
Llama 3.2 3B (Groq)
Tiny Llama on Groq for the cheapest routing and classification paths. Use when 8B Instant is still more model than you need.
Input $0.060 / Output $0.060 per 1M
Context 128K · Approx
Quick take
- Llama 3.2 1B (Groq) is cheaper on input in this catalog snapshot.
- Llama 3.2 1B (Groq) is cheaper on output in this catalog snapshot.
- Both last verified on their model pages (2026-08-02 / 2026-08-02).
Rate table
| Metric | Llama 3.2 1B (Groq) | Llama 3.2 3B (Groq) |
|---|---|---|
| Input / 1M | $0.040 | $0.060 |
| Output / 1M | $0.040 | $0.060 |
| Cached input / 1M | n/a | n/a |
| Context | 128K | 128K |
| Tokenizer | Approx | Approx |