Llama 3.1 8B Instant
$0.05
Input / MTok
$0.08
Output / MTok
Hosted third-party inference; price is Groq's serverless rate, not Meta's own API.
API identifier llama-3.1-8b-instant · Context 131,072 tokens · Max output 131,072
· vendor pricing page
Full price sheet
| Metric | USD / million tokens |
|---|---|
| Input | $0.05 |
| Output | $0.08 |
Price history
| Observed | Metric | Price |
|---|---|---|
| 2026-08-07 19:44:55 | Input | $0.05 |
| 2026-08-07 19:44:55 | Output | $0.08 |
Current prices are free and always will be. The full series — every observation since we started watching, which nobody can reconstruct after the fact — is the part Pro pays for.
Changes affecting this model
coverage
Llama 3.1 8B Instant added to coverage
—
2026-08-07Compare with
Models at a similar price point.
| Model | Input | Output | |
|---|---|---|---|
| LFM2.5 8B-A1B | $0.03 | $0.12 | compare → |
| Ministral 3 3B | $0.1 | $0.1 | compare → |
| Gemma 3n E4B IT | $0.06 | $0.12 | compare → |
| Ternary Bonsai 27B | $0 | $0 | compare → |
| Ministral 3 8B | $0.15 | $0.15 | compare → |
| GPT-OSS 20B | $0.05 | $0.2 | compare → |