Head to head
Llama 3.3 70B Instruct Turbo vs Gemini 3.1 Flash-Lite
On a blended workload of one million input tokens and three million output tokens, Llama 3.3 70B Instruct Turbo is 1.1× cheaper than Gemini 3.1 Flash-Lite.
| Llama 3.3 70B Instruct Turbo | Gemini 3.1 Flash-Lite | |
|---|---|---|
| Provider | Together AI | |
| Input / MTok | $1.04 | $0.25 |
| Output / MTok | $1.04 | $1.5 |
| Cache read / MTok | — | $0.025 |
| Context window | 131,072 | — |
| Blended 1M in + 3M out | $4.16 | $4.75 |
| Status | ga | ga |
Llama 3.3 70B Instruct Turbo detail Gemini 3.1 Flash-Lite detail