Head to head

o3 vs Qwen3.8 2.4T-A95B

On a blended workload of one million input tokens and three million output tokens, Qwen3.8 2.4T-A95B is 1.2× cheaper than o3.

o3Qwen3.8 2.4T-A95B
Provider OpenAI Together AI
Input / MTok $2 $2.5
Output / MTok $8 $6.25
Cache read / MTok $0.5
Context window 200,000
Blended 1M in + 3M out $26.00 $21.25
Status ga ga

o3 detail   Qwen3.8 2.4T-A95B detail