Head to head

Qwen2.5 7B Instruct Turbo vs DeepSeek V4 Flash (0731)

On a blended workload of one million input tokens and three million output tokens, DeepSeek V4 Flash (0731) is 1.2× cheaper than Qwen2.5 7B Instruct Turbo.

Qwen2.5 7B Instruct TurboDeepSeek V4 Flash (0731)
Provider Together AI Together AI
Input / MTok $0.3 $0.14
Output / MTok $0.3 $0.28
Cache read / MTok
Context window 32,768 1,000,000
Blended 1M in + 3M out $1.20 $0.98
Status ga ga

Qwen2.5 7B Instruct Turbo detail   DeepSeek V4 Flash (0731) detail