Head to head

Inkling Small vs Llama 3.3 70B Instruct Turbo

On a blended workload of one million input tokens and three million output tokens, Inkling Small is line-ball with Llama 3.3 70B Instruct Turbo.

Inkling SmallLlama 3.3 70B Instruct Turbo
Provider Together AI Together AI
Input / MTok $0.5 $1.04
Output / MTok $1.2 $1.04
Cache read / MTok
Context window 524,288 131,072
Blended 1M in + 3M out $4.10 $4.16
Status ga ga

Inkling Small detail   Llama 3.3 70B Instruct Turbo detail