Head to head

Llama 3.3 70B Instruct Turbo vs Gemini 3.1 Flash-Lite

On a blended workload of one million input tokens and three million output tokens, Llama 3.3 70B Instruct Turbo is 1.1× cheaper than Gemini 3.1 Flash-Lite.

Llama 3.3 70B Instruct TurboGemini 3.1 Flash-Lite
Provider Together AI Google
Input / MTok $1.04 $0.25
Output / MTok $1.04 $1.5
Cache read / MTok $0.025
Context window 131,072
Blended 1M in + 3M out $4.16 $4.75
Status ga ga

Llama 3.3 70B Instruct Turbo detail   Gemini 3.1 Flash-Lite detail