token.appMeta › Llama 3.2 3B Instruct

Llama 3.2 3B Instruct pricing & benchmarks

Llama 3.2 3B Instruct is available from Meta at $0.050 per million input tokens and $0.330 per million output tokens ($0.120 blended at 3:1). It accepts up to 131,072 tokens of context. Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it.…

Pricing
Per million tokens · OpenRouter
Input$0.050
Output$0.330
Blended (3:1)$0.120
Specification
meta-llama/llama-3.2-3b-instruct
Context window131K
Max output131K
ReleasedSep 2024
LicenceOpen weights
Modalitiesin:text, out:text
No independently-run benchmark scores are published for Llama 3.2 3B Instruct yet. We show nothing rather than reprinting an unverified figure.
Where to run Llama 3.2 3B Instruct
2 hosts · all priced within ~5% of each other. Blended at 3:1. ⚠ marks an option serving less context than the best available, or re-pricing above a prompt-length threshold. Bracketed labels are the provider's own service tier or region.
HostBlended $/1MIn / OutContextWeightsUptime 24h
Parasail $0.120 $0.050 / $0.330 131K bf16 99.9%
Cloudflare
⚠ 80K context, not 131K
$0.122 $0.051 / $0.335 80K 100.0%
Host prices and uptime from OpenRouter, synced 2026-08-08. Lower-precision weights (fp4, fp8) can cost less and score worse than the same model at bf16 — the score above is not host-specific.
Compare Llama 3.2 3B Instruct
FAQ
How much does Llama 3.2 3B Instruct cost?
Llama 3.2 3B Instruct costs $0.050 per million input tokens and $0.330 per million output tokens on OpenRouter. At a 3:1 input:output mix that blends to $0.120 per million tokens.
What is Llama 3.2 3B Instruct's context window?
Llama 3.2 3B Instruct accepts up to 131,072 tokens of context, and can generate up to 131,072 output tokens.