Skip to content

Meta Llama 4 Scout API pricing

Token prices on OpenRouter and at the provider’s list price, and what your own workload would cost per month.

Input
$0.10
Output
$0.30
Cached input
—

OpenRouter, per 1M tokens. A dash means no listed price.

Your monthly cost

Your workload

90M input and 18M output tokens a month (30-day month; token counts are per request).

Assumes: 1,000 conversations a day, each sending about 3,000 tokens of history and instructions and receiving about 600 tokens of replies.

Input not in cache
$9.00
Output
$5.40
Total per month
$14.40
Per day
$0.48
Per request
$0.00048
Unit
Currency

No cache-read price is published for this model, so a cached share is billed as normal input. Compare this bill with other models.

Every price, with its source

OpenRouter price, per 1M tokens

Input
$0.10
Output
$0.30
Cached input (read)
not listed
Context window
1,310,720 tokens
Maximum output
16,384 tokens
Accepts
text, image
Reasoning mode
no
Listed since
5 Apr 2025

No provider list price is recorded for this model: the figures above are what OpenRouter charges, which can differ from buying directly.

Collected from the OpenRouter models API (466 entries). 99 models are cross-checked against provider list prices in the LiteLLM price table.

About Llama 4 Scout

Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B. It supports native multimodal input...

Description supplied by OpenRouter. Model id for the API: meta-llama/llama-4-scout

Llama 4 Scout against other models

Other Meta models

Questions and answers

How much does Meta Llama 4 Scout cost per million tokens?

At the collection of 3 Oct 2026, 08:40 UTC, OpenRouter listed Meta Llama 4 Scout at $0.10 per million input tokens and $0.30 per million output tokens.

What would Meta Llama 4 Scout cost per month?

For 1,000 requests a day with 3,000 input and 600 output tokens each, over a 30-day month and with no caching, the bill is $14.40. Change the numbers in the simulator on this page to match your own traffic.

Does Meta Llama 4 Scout have a cheaper price for cached input?

The sources list no cache-read price for this model, so this site applies no caching discount to it.

What is the context window of Meta Llama 4 Scout?

1,310,720 tokens, with a maximum output of 16,384 tokens, as reported by OpenRouter.