All AI models · LLM benchmarks · Methodology

Meta

Llama 3.2 11B Instruct

Open Llama instruction model for multilingual chat, reasoning, and coding

Model facts

Context window
128K tokens
Maximum output
128K tokens
Cheapest paid input
$0.07 per 1M tokens
Cheapest paid output
$0.33 per 1M tokens
Open weights
yes
Providers
2

Capabilities and modalities

structured output, temperature control, open weights, text.

Artificial Analysis benchmarks

Compare this model on the full LLM benchmark leaderboard.

ConfigurationAAICodingMathOutput tokens/s
Vision3.0not measured1.724.0

Observed price history

  • 2026-08-24: $0.07 input / $0.33 output per 1M tokens

API providers

  • DevPass (LLM Gateway) — model id llama-3.2-11b-instruct; $0.07 input / $0.33 output per 1M tokens; provider documentation
  • LLM Gateway (Inference.net) — model id inference.net/llama-3.2-11b-instruct; $0.07 input / $0.33 output per 1M tokens; provider documentation