Head to head
Inkling Small vs Llama 3.3 70B Instruct Turbo
On a blended workload of one million input tokens and three million output tokens, Inkling Small is line-ball with Llama 3.3 70B Instruct Turbo.
| Inkling Small | Llama 3.3 70B Instruct Turbo | |
|---|---|---|
| Provider | Together AI | Together AI |
| Input / MTok | $0.5 | $1.04 |
| Output / MTok | $1.2 | $1.04 |
| Cache read / MTok | — | — |
| Context window | 524,288 | 131,072 |
| Blended 1M in + 3M out | $4.10 | $4.16 |
| Status | ga | ga |