Get the app

Llama 3.2 1B Instruct: price, context window and capabilities

Llama 3.2 1B Instruct is a model from Meta (meta-llama/llama-3.2-1b-instruct). It does not list tool support. It costs $0.027 per million input tokens and $0.201 per million output tokens, with a 60K-token context window.

$0.027input / 1M tokens
$0.201output / 1M tokens
60Kcontext window
textaccepts

What it costs in practice

Using the listed prices: a chat turn of 2,000 input and 500 output tokens costs about $0.0002, so 10,000 such turns cost about $1.54. Summarising a 100,000-token document into 2,000 tokens costs about $0.0031. A long agent run that reads 1 million tokens and writes 50,000 costs about $0.037.

How it compares

Among the 40 models we track, Llama 3.2 1B Instruct ranks #3 by input price and #2 by output price (1 is cheapest), and #40 by context window (1 is largest). No tracked model with the same tool and reasoning support is cheaper per output token. Models with a larger context window at the same or lower output price: Ministral 3 14B 2512 (262K).

More from Meta

Prices and limits come from OpenRouter's public model catalogue, last synced 4 Oct 2026. Provider pricing can differ from OpenRouter's and changes often; confirm on the provider's own page before you commit. See the guide to choosing a model or the full comparison.

The daily AI brief, on your phone.

Feed, daily deep-dive and bytes — readable offline, with push alerts for the topics you follow.

Get it on Google Play