Skip to content

Meta Llama 3.2 3B Instruct API pricing

Token prices on OpenRouter and at the provider’s list price, and what your own workload would cost per month.

Input
$0.05
Output
$0.33
Cached input
—

OpenRouter, per 1M tokens. A dash means no listed price.

Your monthly cost

Your workload

90M input and 18M output tokens a month (30-day month; token counts are per request).

Assumes: 1,000 conversations a day, each sending about 3,000 tokens of history and instructions and receiving about 600 tokens of replies.

Input not in cache
$4.50
Output
$5.94
Total per month
$10.44
Per day
$0.348
Per request
$0.00035
Unit
Currency

No cache-read price is published for this model, so a cached share is billed as normal input. Compare this bill with other models.

Every price, with its source

OpenRouter price, per 1M tokens

Input
$0.05
Output
$0.33
Cached input (read)
not listed
Context window
131,072 tokens
Maximum output
117,964 tokens
Accepts
text
Reasoning mode
no
Listed since
25 Sep 2024

No provider list price is recorded for this model: the figures above are what OpenRouter charges, which can differ from buying directly.

Collected from the OpenRouter models API (466 entries). 99 models are cross-checked against provider list prices in the LiteLLM price table.

About Llama 3.2 3B Instruct

Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it...

Description supplied by OpenRouter. Model id for the API: meta-llama/llama-3.2-3b-instruct

Other Meta models

Questions and answers

How much does Meta Llama 3.2 3B Instruct cost per million tokens?

At the collection of 3 Oct 2026, 08:40 UTC, OpenRouter listed Meta Llama 3.2 3B Instruct at $0.05 per million input tokens and $0.33 per million output tokens.

What would Meta Llama 3.2 3B Instruct cost per month?

For 1,000 requests a day with 3,000 input and 600 output tokens each, over a 30-day month and with no caching, the bill is $10.44. Change the numbers in the simulator on this page to match your own traffic.

Does Meta Llama 3.2 3B Instruct have a cheaper price for cached input?

The sources list no cache-read price for this model, so this site applies no caching discount to it.

What is the context window of Meta Llama 3.2 3B Instruct?

131,072 tokens, with a maximum output of 117,964 tokens, as reported by OpenRouter.