Meta Llama 4 Scout API pricing
Token prices on OpenRouter and at the provider’s list price, and what your own workload would cost per month.
- Input
- $0.10
- Output
- $0.30
- Cached input
- —
OpenRouter, per 1M tokens. A dash means no listed price.
Your monthly cost
- Input not in cache
- $9.00
- Output
- $5.40
- Total per month
- $14.40
- Per day
- $0.48
- Per request
- $0.00048
No cache-read price is published for this model, so a cached share is billed as normal input. Compare this bill with other models.
Every price, with its source
OpenRouter price, per 1M tokens
- Input
- $0.10
- Output
- $0.30
- Cached input (read)
- not listed
- Context window
- 1,310,720 tokens
- Maximum output
- 16,384 tokens
- Accepts
- text, image
- Reasoning mode
- no
- Listed since
- 5 Apr 2025
No provider list price is recorded for this model: the figures above are what OpenRouter charges, which can differ from buying directly.
Collected from the OpenRouter models API (466 entries). 99 models are cross-checked against provider list prices in the LiteLLM price table.
About Llama 4 Scout
Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B. It supports native multimodal input...
Description supplied by OpenRouter. Model id for the API: meta-llama/llama-4-scout
Llama 4 Scout against other models
Other Meta models
Questions and answers
How much does Meta Llama 4 Scout cost per million tokens?
At the collection of 3 Oct 2026, 08:40 UTC, OpenRouter listed Meta Llama 4 Scout at $0.10 per million input tokens and $0.30 per million output tokens.
What would Meta Llama 4 Scout cost per month?
For 1,000 requests a day with 3,000 input and 600 output tokens each, over a 30-day month and with no caching, the bill is $14.40. Change the numbers in the simulator on this page to match your own traffic.
Does Meta Llama 4 Scout have a cheaper price for cached input?
The sources list no cache-read price for this model, so this site applies no caching discount to it.
What is the context window of Meta Llama 4 Scout?
1,310,720 tokens, with a maximum output of 16,384 tokens, as reported by OpenRouter.
Keep going
- Price tableEvery model and its price, sorted by your monthly bill.
- Cost calculatorYour monthly bill on up to four models, line by line.
- Cheapest modelsThe five lowest-cost models in each category.
- Free modelsModels listed at zero price, kept apart from the paid ones.
- Price changesA dated log of every price that went up or down.
- Caching and batch discountsHow the two discounts work and what you save.
- Long-context pricingModels that read very long prompts, and what they charge.