Head to head

Qwen2.5 7B Instruct Turbo vs Gemini 2.5 Flash-Lite

On a blended workload of one million input tokens and three million output tokens, Qwen2.5 7B Instruct Turbo is 1.1× cheaper than Gemini 2.5 Flash-Lite.

Qwen2.5 7B Instruct TurboGemini 2.5 Flash-Lite
Provider Together AI Google
Input / MTok $0.3 $0.1
Output / MTok $0.3 $0.4
Cache read / MTok $0.01
Context window 32,768
Blended 1M in + 3M out $1.20 $1.30
Status ga ga

Qwen2.5 7B Instruct Turbo detail   Gemini 2.5 Flash-Lite detail