Head to head
GPT-4.1 nano vs Qwen2.5 7B Instruct Turbo
On a blended workload of one million input tokens and three million output tokens, Qwen2.5 7B Instruct Turbo is 1.1× cheaper than GPT-4.1 nano.
| GPT-4.1 nano | Qwen2.5 7B Instruct Turbo | |
|---|---|---|
| Provider | OpenAI | Together AI |
| Input / MTok | $0.1 | $0.3 |
| Output / MTok | $0.4 | $0.3 |
| Cache read / MTok | $0.025 | — |
| Context window | 1,047,576 | 32,768 |
| Blended 1M in + 3M out | $1.30 | $1.20 |
| Status | ga | ga |