Llama 3.2 3B Instruct pricing & benchmarks
Llama 3.2 3B Instruct is available from Meta at $0.050 per million input tokens and $0.330 per million output tokens ($0.120 blended at 3:1). It accepts up to 131,072 tokens of context. Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it.…
Pricing
Per million tokens · OpenRouter
Input$0.050
Output$0.330
Blended (3:1)$0.120
Specification
meta-llama/llama-3.2-3b-instruct
Context window131K
Max output131K
ReleasedSep 2024
LicenceOpen weights
Modalitiesin:text, out:text
No independently-run benchmark scores are published for Llama 3.2 3B Instruct yet. We show nothing rather than reprinting an unverified figure.
Where to run Llama 3.2 3B Instruct
2 hosts · all priced within ~5% of each other. Blended at 3:1. ⚠ marks an option serving less context than the best available, or re-pricing above a prompt-length threshold. Bracketed labels are the provider's own service tier or region.
| Host | Blended $/1M | In / Out | Context | Weights | Uptime 24h |
|---|---|---|---|---|---|
| Parasail | $0.120 | $0.050 / $0.330 | 131K | bf16 | 99.9% |
| Cloudflare ⚠ 80K context, not 131K |
$0.122 | $0.051 / $0.335 | 80K | — | 100.0% |
Host prices and uptime from OpenRouter, synced 2026-08-08. Lower-precision weights (fp4, fp8) can cost less and score worse than the same model at bf16 — the score above is not host-specific.
Compare Llama 3.2 3B Instruct
Llama 3.2 3B Instruct vs Qwen3.8 Max
Llama 3.2 3B Instruct vs DeepSeek V4 Flash 0731
Llama 3.2 3B Instruct vs Claude Opus 5
Llama 3.2 3B Instruct vs Gemini 3.6 Flash
Llama 3.2 3B Instruct vs Kimi K3
Llama 3.2 3B Instruct vs GPT-5.6 Luna
Llama 3.2 3B Instruct vs GPT-5.6 Terra
Llama 3.2 3B Instruct vs GPT-5.6 Sol
FAQ
How much does Llama 3.2 3B Instruct cost?
Llama 3.2 3B Instruct costs $0.050 per million input tokens and $0.330 per million output tokens on OpenRouter. At a 3:1 input:output mix that blends to $0.120 per million tokens.
What is Llama 3.2 3B Instruct's context window?
Llama 3.2 3B Instruct accepts up to 131,072 tokens of context, and can generate up to 131,072 output tokens.