Skip to content

Meta Llama 3.2 1B Instruct API pricing

Token prices on OpenRouter and at the provider’s list price, and what your own workload would cost per month.

Input
$0.027
Output
$0.201
Cached input
—

OpenRouter, per 1M tokens. A dash means no listed price.

Your monthly cost

Your workload

90M input and 18M output tokens a month (30-day month; token counts are per request).

Assumes: 1,000 conversations a day, each sending about 3,000 tokens of history and instructions and receiving about 600 tokens of replies.

Input not in cache
$2.43
Output
$3.62
Total per month
$6.05
Per day
$0.202
Per request
$0.0002
Unit
Currency

No cache-read price is published for this model, so a cached share is billed as normal input. Compare this bill with other models.

Every price, with its source

OpenRouter price, per 1M tokens

Input
$0.027
Output
$0.201
Cached input (read)
not listed
Context window
60,000 tokens
Maximum output
54,000 tokens
Accepts
text
Reasoning mode
no
Listed since
25 Sep 2024

No provider list price is recorded for this model: the figures above are what OpenRouter charges, which can differ from buying directly.

Collected from the OpenRouter models API (466 entries). 99 models are cross-checked against provider list prices in the LiteLLM price table.

About Llama 3.2 1B Instruct

Llama 3.2 1B is a 1-billion-parameter language model focused on efficiently performing natural language tasks, such as summarization, dialogue, and multilingual text analysis. Its smaller size allows it to operate...

Description supplied by OpenRouter. Model id for the API: meta-llama/llama-3.2-1b-instruct

Other Meta models

Questions and answers

How much does Meta Llama 3.2 1B Instruct cost per million tokens?

At the collection of 3 Oct 2026, 08:40 UTC, OpenRouter listed Meta Llama 3.2 1B Instruct at $0.027 per million input tokens and $0.201 per million output tokens.

What would Meta Llama 3.2 1B Instruct cost per month?

For 1,000 requests a day with 3,000 input and 600 output tokens each, over a 30-day month and with no caching, the bill is $6.05. Change the numbers in the simulator on this page to match your own traffic.

Does Meta Llama 3.2 1B Instruct have a cheaper price for cached input?

The sources list no cache-read price for this model, so this site applies no caching discount to it.

What is the context window of Meta Llama 3.2 1B Instruct?

60,000 tokens, with a maximum output of 54,000 tokens, as reported by OpenRouter.