Side by side pricing and context for Llama 3.1 70B (Fireworks) and Llama 3.1 70B (Together). Rates are curated catalog values with last verified dates on each model page.
Llama 3.1 70B (Fireworks)
Llama 3.1 70B on Fireworks for fast open-weight hosting. Compare with Together and Groq 70B-class options on latency and price.
Input $0.900 / Output $0.900 per 1M
Context 128K · Approx
Llama 3.1 70B (Together)
Llama 3.1 70B on Together for open-weight production chat when you want the 3.1 generation instead of 3.3.
Input $0.880 / Output $0.880 per 1M
Context 128K · Approx
Quick take
- Llama 3.1 70B (Together) is cheaper on input in this catalog snapshot.
- Llama 3.1 70B (Together) is cheaper on output in this catalog snapshot.
- Both last verified on their model pages (2026-08-02 / 2026-08-02).
Rate table
| Metric | Llama 3.1 70B (Fireworks) | Llama 3.1 70B (Together) |
|---|---|---|
| Input / 1M | $0.900 | $0.880 |
| Output / 1M | $0.900 | $0.880 |
| Cached input / 1M | n/a | n/a |
| Context | 128K | 128K |
| Tokenizer | Approx | Approx |