Head to head

Llama 3.3 70B Instruct Turbo vs MiniMax M3

On a blended workload of one million input tokens and three million output tokens, MiniMax M3 is 1.1× cheaper than Llama 3.3 70B Instruct Turbo.

Llama 3.3 70B Instruct TurboMiniMax M3
Provider Together AI Together AI
Input / MTok $1.04 $0.3
Output / MTok $1.04 $1.2
Cache read / MTok
Context window 131,072 524,288
Blended 1M in + 3M out $4.16 $3.90
Status ga ga

Llama 3.3 70B Instruct Turbo detail   MiniMax M3 detail