Head to head

Llama 3.1 8B Instant vs Gemma 3n E4B IT

On a blended workload of one million input tokens and three million output tokens, Llama 3.1 8B Instant is 1.4× cheaper than Gemma 3n E4B IT.

Llama 3.1 8B InstantGemma 3n E4B IT
Provider Groq Together AI
Input / MTok $0.05 $0.06
Output / MTok $0.08 $0.12
Cache read / MTok
Context window 131,072 32,768
Blended 1M in + 3M out $0.29 $0.42
Status ga ga

Llama 3.1 8B Instant detail   Gemma 3n E4B IT detail