Llama 4 Scout pricing in 2026
What one task costs on Llama 4 Scout
Per-million-token prices are hard to feel, so here are typical token counts for common jobs priced with Llama 4 Scout. Figures are USD at the rates above and exclude prompt caching and batch discounts.
| Task | Tokens (in / out) | Cost |
|---|---|---|
| Code review of a 500-line pull request | 60,000 / 4,000 | $0.0072 |
| Debug one failing test | 40,000 / 6,000 | $0.0058 |
| Generate a landing page in one shot | 3,000 / 30,000 | $0.0093 |
| Draft a 1,200-word article | 2,000 / 2,500 | $0.0009 |
| Answer a question from a 20-page document | 25,000 / 800 | $0.0027 |
| Multi-step agent run (10 tool calls) | 250,000 / 25,000 | $0.033 |
| All of the above, once | $0.058 |
Cheaper at the same or better capability: Nemotron 3.5 Lightning runs the same set of tasks for $0.040 instead of $0.058.
Llama 4 Scout by Meta (Llama) costs $0.10 per 1M input tokens and $0.30 per 1M output tokens — about 4.0M tokens per $1. Prices last verified 2026-09-19.
The verdict
The largest context window on this site at 1.3M tokens, for $0.10 / $0.30. Its index of 6 means it's a bulk-retrieval and cheap-inference specialist rather than a reasoning model — great for feeding huge documents in, not for hard tasks.
A concrete example: sending 100,000 input tokens and getting back 20,000 output tokens with Llama 4 Scout costs about $0.02 at current prices.
Who is Llama 4 Scout for?
Scout has the largest context window on this site at 1.3M tokens, for about $0.10 / $0.30 per 1M. It's a bulk-retrieval and cheap-inference specialist rather than a reasoning model.
It's the right tool for feeding enormous documents in and getting fast, cheap answers out — not for hard reasoning. For quality work you'll pair it with a stronger model; for raw long-context throughput it's nearly unmatched on price.
Strengths & weaknesses
- Strength: 1.3M context at very low cost
Llama 4 Scout key facts
| Provider | Llama pricing |
|---|---|
| Input price | $0.10 / 1M tokens |
| Output price | $0.30 / 1M tokens |
| Tokens / $1 | 4.0M |
| Intelligence / $1 | 240,000 |
| Context window | 1.3M tokens |
| Tier | fast |
| Intelligence index | 6 / 100 |
| Released | 2025-04-05 |
| Last verified | 2026-09-19 |
What Llama 4 Scout delivers for your budget
For a coding workload, here is roughly what each monthly budget buys:
| Budget | Tokens / month | Months of workload |
|---|---|---|
| $20 | 136.8M | 210.5x |
| $50 | 342.1M | 526.3x |
| $100 | 684.2M | 1052.6x |
Other Llama models
How Llama 4 Scout compares to its closest rivals
Rather than repeating the full catalog on every page, here are the three current models nearest to Llama 4 Scout on the intelligence index — and why you might still pick one instead:
- Llama 4 Maverick (Llama) — $0.19 input / $0.65 output per 1M tokens. Why you might pick it: Open weights you can self-host.
- Mistral Large 3 (Mistral) — $0.50 input / $1.50 output per 1M tokens. Why you might pick it: European open weights.
- Nemotron 3 Super (Nemotron) — $0.08 input / $0.45 output per 1M tokens. Why you might pick it: Near-free input pricing.
Llama 4 Scout — frequently asked questions
How much does Llama 4 Scout cost?
Llama 4 Scout by Meta (Llama) costs $0.10 per 1M input tokens and $0.30 per 1M output tokens.
How many tokens per dollar does Llama 4 Scout give you?
On a typical 1:3 input-to-output mix, Llama 4 Scout delivers around 4.0M tokens per dollar — an intelligence-per-dollar score of 240000.
What is Llama 4 Scout's context window?
Llama 4 Scout supports a context window of 1.3M tokens.
Is Llama 4 Scout worth it in 2026?
Llama 4 Scout is Llama's fast tier model. Use the AI Intelligence Calculator to see how its value compares to every other model at your budget.