Head to head

Gemma 3n E4B IT vs Llama 3.1 8B Instant

On a blended workload of one million input tokens and three million output tokens, Llama 3.1 8B Instant is 1.4× cheaper than Gemma 3n E4B IT.

Gemma 3n E4B ITLlama 3.1 8B Instant
Provider Together AI Groq
Input / MTok $0.06 $0.05
Output / MTok $0.12 $0.08
Cache read / MTok
Context window 32,768 131,072
Blended 1M in + 3M out $0.42 $0.29
Status ga ga

Gemma 3n E4B IT detail   Llama 3.1 8B Instant detail