Head to head
o3 vs Qwen3.8 2.4T-A95B
On a blended workload of one million input tokens and three million output tokens, Qwen3.8 2.4T-A95B is 1.2× cheaper than o3.
| o3 | Qwen3.8 2.4T-A95B | |
|---|---|---|
| Provider | OpenAI | Together AI |
| Input / MTok | $2 | $2.5 |
| Output / MTok | $8 | $6.25 |
| Cache read / MTok | $0.5 | — |
| Context window | 200,000 | — |
| Blended 1M in + 3M out | $26.00 | $21.25 |
| Status | ga | ga |