Llama 3.2 3B (Groq) vs Llama 3.2 3B (Together)

Side by side pricing and context for Llama 3.2 3B (Groq) and Llama 3.2 3B (Together). Rates are curated catalog values with last verified dates on each model page.

Llama 3.2 3B (Groq)

Tiny Llama on Groq for the cheapest routing and classification paths. Use when 8B Instant is still more model than you need.

Input $0.060 / Output $0.060 per 1M

Context 128K · Approx

Llama 3.2 3B (Together)

Tiny Llama on Together for cheap OpenAI-compatible routing. Compare Groq 3B Instant-class hosts when latency is the main constraint.

Input $0.060 / Output $0.060 per 1M

Context 128K · Approx


Quick take

  • Input rates match on this snapshot.
  • Output rates match on this snapshot.
  • Both last verified on their model pages (2026-08-02 / 2026-08-02).

Rate table

MetricLlama 3.2 3B (Groq)Llama 3.2 3B (Together)
Input / 1M$0.060$0.060
Output / 1M$0.060$0.060
Cached input / 1Mn/an/a
Context128K128K
TokenizerApproxApprox