Meta Llama 3.2 3B Instruct API pricing
Token prices on OpenRouter and at the provider’s list price, and what your own workload would cost per month.
- Input
- $0.05
- Output
- $0.33
- Cached input
- —
OpenRouter, per 1M tokens. A dash means no listed price.
Your monthly cost
- Input not in cache
- $4.50
- Output
- $5.94
- Total per month
- $10.44
- Per day
- $0.348
- Per request
- $0.00035
No cache-read price is published for this model, so a cached share is billed as normal input. Compare this bill with other models.
Every price, with its source
OpenRouter price, per 1M tokens
- Input
- $0.05
- Output
- $0.33
- Cached input (read)
- not listed
- Context window
- 131,072 tokens
- Maximum output
- 117,964 tokens
- Accepts
- text
- Reasoning mode
- no
- Listed since
- 25 Sep 2024
No provider list price is recorded for this model: the figures above are what OpenRouter charges, which can differ from buying directly.
Collected from the OpenRouter models API (466 entries). 99 models are cross-checked against provider list prices in the LiteLLM price table.
About Llama 3.2 3B Instruct
Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it...
Description supplied by OpenRouter. Model id for the API: meta-llama/llama-3.2-3b-instruct
Other Meta models
Questions and answers
How much does Meta Llama 3.2 3B Instruct cost per million tokens?
At the collection of 3 Oct 2026, 08:40 UTC, OpenRouter listed Meta Llama 3.2 3B Instruct at $0.05 per million input tokens and $0.33 per million output tokens.
What would Meta Llama 3.2 3B Instruct cost per month?
For 1,000 requests a day with 3,000 input and 600 output tokens each, over a 30-day month and with no caching, the bill is $10.44. Change the numbers in the simulator on this page to match your own traffic.
Does Meta Llama 3.2 3B Instruct have a cheaper price for cached input?
The sources list no cache-read price for this model, so this site applies no caching discount to it.
What is the context window of Meta Llama 3.2 3B Instruct?
131,072 tokens, with a maximum output of 117,964 tokens, as reported by OpenRouter.
Keep going
- Price tableEvery model and its price, sorted by your monthly bill.
- Cost calculatorYour monthly bill on up to four models, line by line.
- Cheapest modelsThe five lowest-cost models in each category.
- Free modelsModels listed at zero price, kept apart from the paid ones.
- Price changesA dated log of every price that went up or down.
- Caching and batch discountsHow the two discounts work and what you save.
- Long-context pricingModels that read very long prompts, and what they charge.