Head to head

Qwen3.8 2.4T-A95B vs GPT-4.1

On a blended workload of one million input tokens and three million output tokens, Qwen3.8 2.4T-A95B is 1.2× cheaper than GPT-4.1.

Qwen3.8 2.4T-A95BGPT-4.1
Provider Together AI OpenAI
Input / MTok $2.5 $2
Output / MTok $6.25 $8
Cache read / MTok $0.5
Context window 1,047,576
Blended 1M in + 3M out $21.25 $26.00
Status ga ga

Qwen3.8 2.4T-A95B detail   GPT-4.1 detail