Meta Llama 3.2 1B Instruct API pricing
Token prices on OpenRouter and at the provider’s list price, and what your own workload would cost per month.
- Input
- $0.027
- Output
- $0.201
- Cached input
- —
OpenRouter, per 1M tokens. A dash means no listed price.
Your monthly cost
- Input not in cache
- $2.43
- Output
- $3.62
- Total per month
- $6.05
- Per day
- $0.202
- Per request
- $0.0002
No cache-read price is published for this model, so a cached share is billed as normal input. Compare this bill with other models.
Every price, with its source
OpenRouter price, per 1M tokens
- Input
- $0.027
- Output
- $0.201
- Cached input (read)
- not listed
- Context window
- 60,000 tokens
- Maximum output
- 54,000 tokens
- Accepts
- text
- Reasoning mode
- no
- Listed since
- 25 Sep 2024
No provider list price is recorded for this model: the figures above are what OpenRouter charges, which can differ from buying directly.
Collected from the OpenRouter models API (466 entries). 99 models are cross-checked against provider list prices in the LiteLLM price table.
About Llama 3.2 1B Instruct
Llama 3.2 1B is a 1-billion-parameter language model focused on efficiently performing natural language tasks, such as summarization, dialogue, and multilingual text analysis. Its smaller size allows it to operate...
Description supplied by OpenRouter. Model id for the API: meta-llama/llama-3.2-1b-instruct
Other Meta models
Questions and answers
How much does Meta Llama 3.2 1B Instruct cost per million tokens?
At the collection of 3 Oct 2026, 08:40 UTC, OpenRouter listed Meta Llama 3.2 1B Instruct at $0.027 per million input tokens and $0.201 per million output tokens.
What would Meta Llama 3.2 1B Instruct cost per month?
For 1,000 requests a day with 3,000 input and 600 output tokens each, over a 30-day month and with no caching, the bill is $6.05. Change the numbers in the simulator on this page to match your own traffic.
Does Meta Llama 3.2 1B Instruct have a cheaper price for cached input?
The sources list no cache-read price for this model, so this site applies no caching discount to it.
What is the context window of Meta Llama 3.2 1B Instruct?
60,000 tokens, with a maximum output of 54,000 tokens, as reported by OpenRouter.
Keep going
- Price tableEvery model and its price, sorted by your monthly bill.
- Cost calculatorYour monthly bill on up to four models, line by line.
- Cheapest modelsThe five lowest-cost models in each category.
- Free modelsModels listed at zero price, kept apart from the paid ones.
- Price changesA dated log of every price that went up or down.
- Caching and batch discountsHow the two discounts work and what you save.
- Long-context pricingModels that read very long prompts, and what they charge.